Full-Time

Deep Learning Software Engineer New Grad

TensorRT Performance

NVIDIA

NVIDIA

10,001+ employees

Designs GPUs and AI HPC platforms

Compensation Overview

$124k - $195.5k/yr

+ Equity

Company Historically Provides H1B Sponsorship

Remote in USA + 1 more

More locations: Santa Clara, CA, USA

Remote

Category
AI & Machine Learning (1)
Software Engineering (1)

Get referred to NVIDIA

See people who can refer or advise you

Requirements
  • Experience with performance analysis and performance optimization
Responsibilities
  • Establish groundbreaking performance benchmarking methodologies and analysis workflows and identify performance issues and opportunities for NVIDIA’s inference ecosystem (e.g. TensorRT/TensorRT-EdgeLLM/Torch-TensorRT)
  • Contribute features and code to NVIDIA/OSS inference frameworks including but not limited to TensorRT/TensorRT-EdgeLLM/Torch-TensorRT.
  • Develop new model pipelines for NVIDIA’s inference ecosystem with optimized performance including but not limited to areas like quantization, scheduling, memory management, and distributed inference to set the gold standard for Gen AI performance.
  • Work with cross-collaborative teams inside and outside of NVIDIA across generative AI, automotive, robotics, image understanding, and speech understanding to set directions and develop innovative inference solutions.
  • Scale performance of deep learning models across different architectures and types of NVIDIA accelerators.
Desired Qualifications
  • Strong foundation and architectural knowledge of GPUs.
  • Deep understanding of modern deep learning models and workloads (e.g. Transformers, Recommenders, ASR, TTS, Visual Understanding).
  • Proficiency in one of the deep learning programming domain specific languages (e.g. CUDA/TileIR/CuTeDSL/cutlass/Triton).
  • Prior contributions to major LLM inference frameworks (e.g. vLLM) or prior experience with graph compilers in deep learning inference (e.g. TorchDynamo/TorchInductor).
  • Prior experience optimizing performance for low-latency, resource-constrained systems or embedded AI pipelines (e.g. Jetson systems or other edge AI accelerators).

NVIDIA designs and manufactures graphics processing units (GPUs) and computing platforms used for gaming, data centers, and artificial intelligence. These products work by using parallel processing to handle complex mathematical calculations much faster than standard computer processors, supported by a software ecosystem that allows developers to build and run AI models. Unlike competitors that may focus solely on hardware, NVIDIA integrates its chips with specialized software and cloud services to create a complete environment for high-performance tasks. The company’s goal is to provide the underlying technology necessary to power advanced computing, from realistic video game graphics to autonomous vehicles and large-scale data analysis.

Company Size

10,001+

Company Stage

IPO

Headquarters

Santa Clara, California

Founded

1993

Get referred to NVIDIA

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • NVIDIA’s February 2026 fiscal revenue reached $215.9 billion, proving unmatched AI demand.
  • August 2026 compute-financing partnerships can unlock hundreds of billions for GPU deployments.
  • India’s Larsen and Toubro AI factory contract broadens NVIDIA’s global infrastructure footprint.

What critics are saying

  • China’s SAMR antitrust probe from December 2024 still threatens sales and remedies in 2026.
  • August 2026 financing platforms create $125 billion backstop exposure if AI buyers default.
  • OpenAI, CoreWeave, and neocloud customers depend on NVIDIA allocation; custom ASICs break pricing power by 2027.

What makes NVIDIA unique

  • CUDA and NVLink lock customers into NVIDIA systems, not commodity accelerators.
  • February 25, 2026 results showed $62.3 billion data center revenue and 75% gross margins.
  • August 10, 2026 financing MOUs with Goldman Sachs, BlackRock, and KKR extend NVIDIA’s reach.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Company Equity

401(k) Company Match

Growth & Insights and Company News

Headcount

6 month growth

0%

1 year growth

-2%

2 year growth

-3%
Yahoo Finance
Aug 16th, 2026
Nvidia's $5B Intel investment yields $25B profit — acquires $21B SpaceX stake

Nvidia's $5 billion Intel stock purchase from last year has generated nearly $25 billion in returns, according to an SEC filing this week. The investment was part of the companies' strategic AI infrastructure partnership announced in September. The filing also revealed Nvidia holds approximately $21 billion in SpaceX stock. The SpaceX investment appears strategic, as xAI has committed to exclusively using Nvidia hardware in its AI data centres. After investing in Intel, Nvidia sold its 1.1 million Arm shares, worth $178.1 million last August. The company continues developing Arm-based CPUs despite divesting its stake. Other significant investments include $2 billion in Coherent for laser technology, $1 billion in Nokia for AI-RAN innovation, and $2 billion in Synopsys for AI-enhanced design tools. All three investments have appreciated in value.

Yahoo Finance
Aug 15th, 2026
Nvidia's $500B AI deal spotlights $70B in hidden tech backstops worrying bond traders

Bond traders are growing concerned about roughly $70 billion in off-balance-sheet liabilities tied to AI company financing. These "residual value" backstops allow major tech firms like Nvidia and Broadcom to support customer debt deals without recording the obligations on their own books. Nvidia's recent $500 billion financing partnership has intensified scrutiny of these arrangements. The structure typically involves a special-purpose vehicle buying chips backed by customer contracts. If buyers default and asset sales fall short, backstoppers like Nvidia cover the difference. Meta Platforms stated in filings that such payments are "not probable" and recorded no liability. Whilst proponents argue chip demand will remain strong and debt gets paid down over time, investors are examining past deals to assess risks as AI chip financing expands rapidly.

Tech in Asia
Aug 15th, 2026
Nvidia scales back data centre guarantee plans amid $500B Stargate AI infrastructure push

Nvidia has scaled back plans to guarantee data centre projects, according to a report. The move comes as OpenAI's Stargate project with Oracle and SoftBank Group advances a $500 billion US AI infrastructure initiative expected to exceed 9 gigawatts by 2029. Seven US sites are under development, with 0.3 gigawatts already operating in Abilene, Texas. Legal analysts noted that large AI infrastructure commitments can involve contingent or off-balance-sheet risk. Stargate's expansion continues, with related sites using bond, private-credit, and special-purpose vehicle financing. Analysts suggest any Nvidia pullback would more likely alter the financing mix rather than halt construction. The report did not detail Nvidia's commitment.

Yahoo Finance
Aug 15th, 2026
Nvidia chips appreciate like 'fine jewelry', says Cramer as CoreWeave surges $17

Jim Cramer highlighted NVIDIA's dominant position in the data centre market on his Mad Money show, citing CoreWeave's strong quarterly results as evidence of the chips' enduring value. CoreWeave CEO Michael Intrator demonstrated that older NVIDIA GPUs retain or even appreciate in value, with some nine-year-old chips remaining highly sought after. NVIDIA reported record fiscal first-quarter revenue of $81.6 billion, up 85% year-over-year. Data centre compute revenue reached $60.4 billion, whilst networking revenue surged 199% to $14.8 billion. CEO Jensen Huang described the global AI infrastructure buildout as unprecedented. Morningstar analyst Brian Colello maintained a fair value estimate of $280 for the stock, acknowledging concerns about complex financing arrangements but affirming underlying chip demand remains strong.

ION Analytics
Aug 14th, 2026
Agility Robotics' $2.5B SPAC debut tests humanoid market at discount to $39B rival valuations

Agility Robotics is going public through a SPAC merger with Churchill Capital Corp XI at a $2.5bn valuation, significantly below private humanoid robotics rivals. The Oregon-based company expects to raise over $620m in proceeds, with the merger closing in Q4 2026. The valuation trails competitors substantially. Apptronik raised funds at above $5bn, whilst Figure AI closed Series C funding at a $39bn post-money valuation. Investors cite Agility's relatively weaker position on deployments and technology as justification for the discount. Agility has booked over $300m in multi-year revenue tied to roughly 1,000 robots, with 65,000 operational hours across nine customer facilities. However, analysts caution this backlog involves contracts for robots still in development, with cancellation provisions. Industry experts warn against overvaluing humanoid robotics relative to established automation technologies. The company's challenge lies in converting technological promise into repeatable deployments and demonstrable ROI whilst competing against proven automation alternatives already generating substantial revenue.