TC

Member of Technical Staff - Performance Modeling

Accepting applications

Touring Capital · San Francisco Bay Area

Full-Time Senior AIC++PythonRTLSoC
Posted
19h ago
Category
Design
Experience
Senior
Country
United States
Infinity Artificial Intelligence Institute San Francisco Bay Area

Member of Technical Staff - Performance Modeling

Infinity Artificial Intelligence Institute San Francisco Bay Area

1 day ago Be among the first 25 applicants

See who Infinity Artificial Intelligence Institute has hired for this role

Join or sign in to save this job

Join to save Member of Technical Staff - Performance Modeling at Infinity Artificial Intelligence Institute

Continue with GoogleContinue with Google

Email or phone

Password

Forgot password?

or

Continue with GoogleContinue with Google

New to LinkedIn? Join now

By clicking Continue to join or sign in, you agree to LinkedIn’s User Agreement, Privacy Policy, and Cookie Policy.

Report this job

Member of Technical Staff - Performance Modeling

Company: Infinity

Team: Systems / AI Infrastructure
Location: San Francisco (on-site)
Type: Full-time

The Mission

Everything upstream of us is gated by one scarce resource: the chip itself. A part is under NDA, or taped out but not yet back from the fab, or there are four boards on earth and all four are booked. That scarcity sets the pace of the entire company, until you take the hardware off the critical path.

That is what the simulator does. It runs on ordinary CPUs and reproduces an AI accelerator’s behavior faithfully enough that Ignition, our bringup agent, can write a compiler backend, and a matmul kernel against it, and Infy, our inference library, can prove a kernel correct, all before the silicon is in the building. Because it runs on CPUs, it scales on infrastructure we already own: hundreds of simulated chips are cheaper, and far more available, than a single reserved board.

This is the substrate the rest of the stack stands on. Every gate in our test ladder runs on real hardware and in simulation at the same time, so a kernel that passes on a chip we have can be trusted on one we don’t yet. The simulator is what lets a brand-new architecture join it months before its first board ships. The same machine has already done the hard version in production.

What you’ll work on

You’ll Own The Models The Rest Of The Stack Builds And Validates Against. Which Of These You Take Depends On Your Strengths

Data-placement and movement tracking. Model exactly what sits in which memory at each step and every transfer between units. This is what elevates the simulator from “runs the ops” to “catches the bugs”: ordering hazards and race conditions that only appear on real silicon under precisely the wrong schedule become reproducible and inspectable here.
Performance and timing models. Capture enough of the memory hierarchy, execution units, and interconnect to predict throughput and quantify the gap to peak with no physical part in hand. Because the model observes every stall and every byte moved, its benchmarks rest on mechanism rather than curve-fitting, and moving them from plausible to trustworthy is where much of the difficulty lives.
Coverage across execution models. Warp-based GPUs, scalar tile meshes, dataflow arrays, flat SIMD, analog MAC, wafer-scale: each demands a different simulator skeleton. The leverage is in the shared abstractions, an execution-model-neutral core a new architecture can slot into instead of forcing a rewrite.
Continuous fidelity checking. Design the reconciliation loop that pins each model to its silicon: bootstrap from a spec or fuzzer output, then, whenever a board is available, diff the simulator against the real chip and drive any divergence to zero before it spreads.
Functional and ISA-level modeling. Take a spec, or more often the behavioral model our probes and ISA fuzzer assemble from a chip with no complete spec, and turn it into an executable that runs the instruction set with correct semantics and can be checked bit-for-bit against a reference. This is the layer everything else trusts.
RTL co-simulation. Stand up Verilator or Icarus Verilog against a vendor’s RTL or a partial model and wire it into the same test harness the physical part uses, so a design can be validated before first silicon exists.
The golden reference. The simulator is the oracle the entire test ladder trusts, which makes its correctness non-negotiable: a wrong reference doesn’t merely fail, it silently certifies broken code as correct. Owning this means owning that standard.
Speed and scale. A model no one can afford to run is a model no one uses. JIT compilation, parallel execution, per-layer fidelity that spends cycles only where they matter, and horizontal scale-out across CPU fleets so that hundreds of simulated chips cost less than one reserved board.
Integration with the agent loop. Expose the same hardware schema, probes, and interfaces the agent sees on real targets, so Ignition cannot tell whether it is driving silicon or a simulation until the moment that distinction actually matters.

What we’re looking for

We weight range and depth over any particular résumé. Strong candidates will have most of the following:

Working computer-architecture intuition. Not a textbook recall of pipelines and memory hierarchies, but a felt sense of what actually determines throughput and where the stalls hide.
You’ve built or extended a simulator or emulator, functional or cycle-accurate: gem5, QEMU, Spike, Verilator, or something you wrote from scratch.
You’re comfortable when the spec is incomplete or wrong. Reverse-engineering real behavior, reconciling a datasheet that disagrees with the silicon, and shipping a model you can defend anyway is the normal case here, not the exception.
Fluency in Python and a systems language: C, C++, or Rust.
A test-first instinct. You understand, without being told, that the thing everything else is measured against has to be held to a higher standard than the code it measures.

Nice to have

Built an ISA simulator or emulator end to end.
A hardware-design background: HLS, Verilog, or time on the vendor side of a chip.
Familiarity with non-GPU execution models; the more exotic, the better.
Performance modeling or roofline analysis as part of your regular practice.
Experience building with coding agents, and calibrated judgment about where the model is trustworthy and where the tests have to catch it.

Who we are

Infinity is an early-stage AI infrastructure research company building the software layer that makes non-NVIDIA chips competitive for AI inference. Rather than relying on scarce human kernel engineers, we use AI to automatically generate, test, and optimize the low-level code that determines how efficiently a chip runs AI models. We’ve signed or are negotiating design partnerships with d-Matrix, AMD, AWS Trainium, Microsoft (Maia and Nexus), Qualcomm, and others. Founded by Jeremy Nixon (former Google Brain; co-founder of AGI House with Andrej Karpathy), Infinity has raised $15M from investors including the founder of Intercom, the VP of AI at AMD, and the founder of MLCommons. We’re headquartered in San Francisco.

Seniority level Mid-Senior level
Employment type Full-time
Job function Engineering and Information Technology
Industries Software Development

Referrals increase your chances of interviewing at Infinity Artificial Intelligence Institute by 2x

See who you know

Get notified about new Member of Technical Staff jobs in San Francisco Bay Area.

Sign in to create job alert

Jobs
Senior Member of Technical Staff jobs
Member of Technical Staff - Performance Modeling

Similar jobs

Performance Modeling Engineer

Performance Modeling Engineer

OpenAI

San Francisco, CA $293,000 - $385,000 2 weeks ago

Performance Modeling Engineer

Performance Modeling Engineer

MediaTek

San Jose, CA 1 day ago

Workload / Performance Model Lead

Workload / Performance Model Lead

SiFive

Berkeley, CA 4 months ago

Principal Performance Modeling Engineer

Principal Performance Modeling Engineer

AMD

Santa Clara, CA $188,160 - $282,240 1 day ago

Performance & Capacity Engineering - Capacity Planning Optimization

Performance & Capacity Engineering - Capacity Planning Optimization

Meta

Menlo Park, CA $184,000 - $257,000 5 hours ago

System Performance Modeling Engineer

System Performance Modeling Engineer

AMD

Santa Clara, CA $189,600 - $284,400 5 days ago

Principal Performance Architect

Principal Performance Architect

Microsoft

Mountain View, CA 1 day ago

Performance Modeling Engineer ~2

Performance Modeling Engineer ~2

OpenAI

San Francisco, CA

$293,000.00

$385,000.00

2 weeks ago

Performance Modeling Engineer

Performance Modeling Engineer

Etched

San Jose, CA

$175,000.00

$275,000.00

1 week ago

CPU Performance Modeling Engineer (Multiple Levels)

CPU Performance Modeling Engineer (Multiple Levels)

Qualcomm

Santa Clara, CA 3 days ago

System Architect

System Architect

Micron Technology

San Jose, CA 1 week ago

Senior Performance Verification Engineer

Senior Performance Verification Engineer

NVIDIA

Santa Clara, CA 2 weeks ago

Senior Performance Architect, Nemotron

Senior Performance Architect, Nemotron

NVIDIA

Santa Clara, CA 3 weeks ago

CPU Performance Analysis Engineer (Multiple Locations- San Diego, Santa Clara, Austin)

CPU Performance Analysis Engineer (Multiple Locations- San Diego, Santa Clara, Austin)

Qualcomm

Santa Clara, CA 4 days ago

CPU Performance Architect

CPU Performance Architect

Google

Mountain View, CA 3 days ago

Systems Modeling & Optimization Engineer

Systems Modeling & Optimization Engineer

Waymo

Mountain View, CA 3 days ago

AI Performance Modeling Engineer

AI Performance Modeling Engineer

Quadric

Burlingame, CA

$150,000.00

$200,000.00

1 week ago

Principal Performance Modeling Engineer

Principal Performance Modeling Engineer

Oho Group

San Francisco Bay Area 3 hours ago

HPC/AI Performance Specialist

HPC/AI Performance Specialist

Berkeley Lab

Berkeley, CA 1 month ago

Member of Technical Staff — Performance Palo Alto, CA

Member of Technical Staff — Performance Palo Alto, CA

RadixArk

Palo Alto, CA 1 month ago

Research Scientist, Infrastructure Modeling and Reliability

Research Scientist, Infrastructure Modeling and Reliability

Meta

Menlo Park, CA

$271,000.00

$347,000.00

2 days ago

Principal Performance Modeling Architect

Principal Performance Modeling Architect

Oxmiq Labs

Campbell, CA 23 hours ago

Member of Technical Staff

Member of Technical Staff

Morph

San Francisco, CA

$175,000.00

$350,000.00

1 week ago

Performance Engineer, Inference Systems

Performance Engineer, Inference Systems

Anthropic

San Francisco, CA 1 hour ago

SoC Performance Architect

SoC Performance Architect

Samsung Semiconductor

San Jose, CA 1 week ago

Senior Power Architect, Power and Performance Analysis Tools

Senior Power Architect, Power and Performance Analysis Tools

NVIDIA AI

Santa Clara, CA 3 days ago

R&D Engineer

R&D Engineer

Bolt Graphics

Sunnyvale, CA 1 week ago

People also viewed

Senior Performance Engineer

Senior Performance Engineer

Samsung Semiconductor

San Jose, CA 1 day ago

High Performance Computing Engineer

High Performance Computing Engineer

SLB

Sunnyvale, CA 5 days ago

Senior Performance Co-Design Engineer, TPU

Senior Performance Co-Design Engineer, TPU

Google

Sunnyvale, CA 2 weeks ago

SystemC Modeling Developer (Remote)

SystemC Modeling Developer (Remote)

UST

San Francisco, CA 1 week ago

Sr. Performance Modeling Architect

Sr. Performance Modeling Architect

Tenstorrent

Austin, CA 6 days ago

Performance Modeling Engineer

Performance Modeling Engineer

Acceler8 Talent

Santa Clara, CA $250,000 - $350,000 6 days ago

Datacenter Compute SoC Perf/Power Modeling Architect

Datacenter Compute SoC Perf/Power Modeling Architect

MediaTek

San Jose, CA 1 week ago

Performance Modeling Lead

Performance Modeling Lead

OpenAI

San Francisco, CA $293,000 - $385,000 2 weeks ago

Senior Power Architect, Power and Performance Analysis Tools

Senior Power Architect, Power and Performance Analysis Tools

NVIDIA

Santa Clara, CA 2 days ago

CPU Performance Research Engineer

CPU Performance Research Engineer

Qualcomm

Santa Clara, CA 6 days ago

Similar Searches

Senior Member of Technical Staff jobs

8,067 open jobs

Member Technical jobs

62,993 open jobs

Senior Wireless Engineer jobs

35,731 open jobs

Physical Design Engineer jobs

6,753 open jobs

Staff Test Engineer jobs

2,677 open jobs

Principal Firmware Engineer jobs

1,780 open jobs

Line Technician jobs

109,890 open jobs

Lead Infrastructure Engineer jobs

16,227 open jobs

Support Team Manager jobs

83,600 open jobs

Vice President Software jobs

49,146 open jobs

Senior Lead Software Engineer jobs

49,381 open jobs

Principal Researcher jobs

4,530 open jobs

Switch Engineer jobs

9,391 open jobs

Staff Software Engineer jobs

64,945 open jobs

Lead Quality Engineer jobs

10,548 open jobs

Control Coordinator jobs

39,876 open jobs

Market Maker jobs

1,432 open jobs

Yield Engineer jobs

9,445 open jobs

Computer Scientist jobs

49,477 open jobs

Lead Test Engineer jobs

13,921 open jobs

House Supervisor jobs

29,485 open jobs

Core Engineer jobs

33,936 open jobs

Cable Technician jobs

11,124 open jobs

Principal Software Engineer jobs

73,845 open jobs

Logic Design Engineer jobs

1,858 open jobs

Explore top content on LinkedIn

Find curated posts and insights for relevant topics all in one place.

View top content
Show more Show less