Senior Sales Engineer - Token Factory
Remote · Full-time
- Posted 1mo ago
- From Nebius’s careers page
- Work mode
- Remote
- Type
- Full-time
- Level
- Senior
- Department
- Sales
Apply on Nebius’s site
Opens the listing on careers.nebius.com
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.
About the role
The role\n\nWe are building a high-performance AI inference platform for developer-native teams running latency- and cost-sensitive workloads at scale.\n\nIn AI infrastructure, PoC success does not guarantee production success. This role ensures that what we commit to is scalable, efficient, and aligned with platform strategy.\n\nWe are looking for a Senior Sales Engineer to become a foundational technical partner to our customers and a force multiplier for Sales and Engineering. You will shape complex AI workloads from first discovery through production feasibility validation, ensuring technical rigor, economic viability, and scalable architecture decisions.\n\nYou will operate at the intersection of customer ambition, engineering reality, and commercial growth, in influencing:\n- Revenue quality\n- Engineering focus\n- Product evolution\n- Customer trust at scale\n\nYou’re welcome to work remotely from Europe.\n\n### Your responsibilities will include:\n- Strategic Technical Discovery\n- Lead deep technical discovery with engineering teams and technical founders\n- Understand model requirements, traffic expectations, latency constraints, GPU economics, and system dependencies\n- Translate customer ambition into production-feasible architectures\n- Identify hidden technical risks early\n\n- Commercial Acceleration\n- Partner tightly with Sales on strategic deals\n- Influence deal strategy through architectural clarity\n- Prevent misaligned commitments before engineering allocation\n- Increase PoC-to-production conversion by ensuring technical realism\n\n- PoC Architecture & Validation\n- Define measurable success criteria (latency, TTFT, throughput, cost envelope)\n- Classify workload complexity and required optimization depth\n- Align appropriate resources (ML Solution Architects, engineering, GPU capacity, etc.)\n- Drive structured Go / No-Go decisions\n- Prevent uncontrolled customization or hidden R&D\n\n- Pattern Recognition & Platform Leverage\n- Identify recurring configuration patterns across customers\n- Quantify demand for advanced optimizations (quantization, speculative decoding, etc.)\n- Surface structured insights to Product and Engineering\n- Help evolve platform capabilities based on real workload data\n\n### We expect you to have:\n- Deep understanding of AI inference systems and GPU-backed infrastructure\n- Experience with LLM workloads and performance-sensitive environments\n- Experience with inference frameworks and libraries (e.g., vLLM, SGLang, TensorRT-LLM).\n- Ability to reason about latency, throughput, cost, and architecture tradeoffs\n- Strong customer presence with engineering-first organizations\n- Comfort challenging assumptions and pushing back constructively\n- Commercial awareness – you understand that engineering time is a strategic resource\n\n### Preferred technical stack:\n- Programming Languages– Python\n- Frameworks and Libraries– vLLM, SGLang, TensorRT-LLM, OpenAI/Anthropic SDKs\n- Frameworks for Agentic Pipelines : Langchain / Langsmith / smolagents / equivalent\n- API and Web Frameworks– FastAPI, Flask\n- MLOps and DevOps tools– Kubernetes (K8s), Docker, Git\n- Cloud Platforms– AWS (SageMaker, Bedrock), GCP (Vertex AI), Azure (Azure ML)\n\n### What success looks like:\n- Strategic deals are technically sound before engineering engagement\n- PoCs are clearly scoped and economically justified\n- Engineering capacity is allocated predictably\n- Conversion to production improves\n- Customers view you as a trusted architectural advisor\n\n### Benefits & Perks:\n- Competitive compensation\n- Career growth and learning opportunities\n- Flexibility and ownership\n- Collaborative and innovative culture\n- Opportunity to work on impactful AI projects\n- International environment and talented teams\n\n### What's it like to work at Nebius:\nFast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI\n\n### Equal Opportunity Statement:\nNebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.\n\nApplicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire.\n\nIf you need accommodations during the application process, please let us know.
Skills they ask for
Pick one to see other roles that ask for it.
About Nebius
Cloud infrastructure for AINebius provides cloud infrastructure and services for building and scaling AI workloads.
See all 133 roles at NebiusMore roles at Nebius
See all 133- Principal ML Solutions Architect - Token FactoryUnited States · Principal · RemoteEngineering · Principal · RemoteUnited States11h
- Delivery Operations Manager - Token FactoryRemoteBusiness Operations · Remote11h
- Head of Strategic Partnerships, TavilyUnited States · Director · RemoteSales · Director · RemoteUnited States16h
- General Manager, Data Center (New Build)Independence · Director · On-siteInformation Technology · Director · On-siteIndependence, United States1d
Share this role
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.