This one is closed
Live roles like this one
-
B
2h ago
Software Architect (AWS/Azure) - Canada
Intelerad Canada CA$94k - CA$125k/yr
-
D
2h ago
Ergomed Spain
- D 7h ago
- C 7h ago
See every "Large Scale Neural" role →
Get new “Large Scale Neural” roles by email
One email a day with what is new in "Large Scale Neural". Nothing new, no email.
We confirm the address first, and every mail carries an unsubscribe link. Alerts are ours, not a third party's.
Why this grade This listing scored 57/100, which is a C. It lost the most ground on pay transparency. See the breakdown
- Description depth 20 / 20 How much the posting actually says about the work, measured in characters of real text.
- Freshness 12 / 15 How recently it was posted. Older postings are likelier to be filled or abandoned.
- Pay transparency 12 / 25 A published salary range, worth more than any other single factor because it is what a candidate cannot find out without applying.
- Remote clarity 8 / 15 Whether "remote" means anywhere, or is quietly restricted to one country.
- Corroboration 5 / 10 Whether more than one source carries this listing.
- Role specificity 0 / 10 Whether the listing is tagged well enough to tell what the role actually is.
Every figure above is arithmetic over the posting itself — its salary field, its text, its age, its tags and how many sources carry it. How the grades work →
This listing does not state a salary
$140k – $226k
That is the middle half of what comparable roles paid on this board over the last 90 days — 31 listings that did publish a figure, median $186k. It is not this employer's offer, and we have no idea what they pay. It is only what the rest of the market advertised.
We are seeking a Senior Solutions Architect with deep expertise in large-scale neural network inference and a proven ability to lead technical collaboration with frontier AI labs and enterprises deploying AI at scale across EMEA. In this role, you will define the technical direction for AI inference across EMEA by identifying critical bottlenecks and driving the development of scalable, high-impact solutions. By aligning key team members within NVIDIA and customer organizations, you will influence strategic technology decisions to develop the deployment of next-generation AI inference at scale.
What you will be doing:
Lead the inference strategy for a portfolio of EMEA AI Natives customers, guiding engagements from initial proof of concept to production-scale deployments.
Identify inference challenges across customer deployments including latency, efficiency, cost per token, memory utilization, and low-latency networking.
Architect and optimize high-performance inference pipelines using NVIDIA Dynamo, TensorRT-LLM, vLLM, SGLang, and other inference backends, improving GPU utilization and AI cluster efficiency.
Translate customer insights and deployment patterns into actionable product feedback that develops the roadmap for NVIDIA stack such as Dynamo, TensorRT-LLM, and NIM.
What we need to see:
MS or PhD in Computer Science, Engineering, or equivalent experience in the field.
8+ years in AI/ML infrastructure, with deep expertise in LLM/VLM inference optimization and production deployment at scale.
Deep understanding of transformer inference acceleration: quantization (INT4/FP8), speculative decoding, disaggregated inference, continuous batching, KV cache optimization, and WideEP for MoE models.
Understanding of GPU memory hierarchies and low-latency networking along with their influence on inference performance.
Proven track record to lead technical initiatives.
Excellent communication skills, effective with research scientists, infrastructure engineers, and executive team members.
Ways to stand out from the crowd:
Experience with NVIDIA's inference stack, including TensorRT-LLM, Triton Inference Server, NIM, and NVIDIA Dynamo.
Experience with GPU orchestration on Kubernetes.
You have operated inference at scale inside a frontier AI lab or hyperscale's inference team.
Contributions to open-source inference projects such as vLLM, SGLang, KServe, or NVIDIA Dynamo.
Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer
NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. For Poland: The base salary range is 292,500 PLN - 507,000 PLN for Level 4, and 375,000 PLN - 650,000 PLN for Level 5.Originally posted on Himalayas
Apply for this role Opens himalayas.app — the link as listed; we have not yet verified it is the employer's own page
Quick question · anonymous · one tap
Would you apply to this job?
Answer to see what other job seekers said.
Your turn · no account needed
Help the next applicant
You may know something about this listing that we cannot see from here. One tap. No account needed. Signed-in reports earn points once the evidence agrees with you.
I know what it pays
Sign in with Google to earn points for reports — 100 confirmed points buy a week of Early Access.
Where this listing came from
- 10 Sep 2026 Himalayas first sighting
Seen on 1 board over 0 days.