Brellium Senior AI Engineer at Brellium owning and evolving the intelligence engine powering the platform. Design and improve LLM-powered pipelines and agent architectures focusing on accuracy, latency, reliability, and cost efficiency at scale.
Responsibilities
We're hiring a Senior AI Engineer to own and evolve the intelligence engine powering Brellium. This role goes well beyond toy demos — it requires deep fluency in prompt design, evaluation frameworks, agentic systems, structured extraction, tool use, retrieval pipelines, and performance tradeoffs. The work is core: transforming raw session data into clinically structured outputs, with a constant focus on improving accuracy, latency, reliability, and cost efficiency at scale.
Design and improve LLM-powered pipelines and agent architectures
Build robust evaluation and benchmarking frameworks to measure output quality and regression
Optimize for latency, cost, determinism, and output quality across high daily volume
Implement retrieval-augmented generation (RAG) and tool-use strategies where appropriate
Ship production-grade AI systems – not experiments
Collaborate closely with backend and product teams to integrate AI capabilities into the core platform
YOU MIGHT BE A FIT IF YOU
Have built and deployed LLM systems in production at meaningful scale
Understand modern agent patterns: tool calling, RAG, memory, orchestration
Think deeply about evaluation, hallucination mitigation, and reliability
Care about performance and operational excellence, not just capability
Qualification
Deep familiarity with LLM concepts
Required
5+ years of professional software engineering experience
Hands-on experience building and shipping LLM-powered applications in production
Deep familiarity with LLM concepts: prompt engineering, structured extraction, function/tool calling, RAG, evaluation frameworks
Experience with AWS or comparable cloud infrastructure for AI workloads
Strong software engineering fundamentals – you write clean, maintainable systems, not just notebooks