inhousefyi
← Back to listings

Senior AI Engineer (Edge Dialog Systems)

BrightAI CorporationPalo Alto, California, United States · Posted 4 days ago
Full-timeEst. 175,000 USD
Apply now

Description

Senior AI Engineer – Edge Dialog Systems

BrightAI is a high-growth Physical AI company transforming how businesses interact with the physical world through intelligent automation. Our AI platform processes visual, spatial, and temporal data from billions of real-world events—captured across edge devices, mobile sensors, and cloud infrastructure—to enable intelligent decision-making at scale.

We are now hiring a Sr. AI Engineer – Edge Dialog Systems to own and evolve the on-device conversational AI that powers our industrial safety wearable. The assistant guides field technicians through safety-critical procedures by voice, and it runs on the device itself, under hard latency, memory, and thermal budgets, with deterministic safeguards that take precedence over model output.

This is a systems role rather than a prompt-and-retrieve role. The competencies of a strong LLM and RAG engineer are needed for the position, but the work itself is dialog systems engineering at the edge, where most conversational turns are resolved by deterministic and embedding-based methods, and the language model is the last resort rather than the first move.

You will work at the intersection of natural language understanding (NLU), small language models (SLM), and embedded software, building a system in which a wrong answer is a safety concern and not merely a quality issue.

Responsibilities

  • Own the on-device dialog pipeline end to end: intent routing, hybrid intent classification (pattern matching combined with embedding similarity and out-of-domain detection), text normalization for noisy speech input, and the multi-step guided-procedure engine.
  • Maintain and extend the deterministic safety layer that wraps the language model—confirmation and echo-back gating, criticality tagging, negation handling—so that a misheard answer on a safety-critical step cannot pass silently.
  • Run SLM inference on-device under memory, computational complexity, and latency budgets, and reduce the per-turn inference cost through model selection, quantization, and runtime optimization.
  • Preserve and extend the zero-shot configuration model, in which new device commands and customer procedures are authored as data rather than released as code, so that a new customer can be onboarded in hours rather than weeks.
  • Coordinate the device deployment pipeline with the edge team
  • Maintain the API contract with on-device voice pipeline & its speech-to-text (STT) stack.
  • Define and run on-device benchmarks: latency, accuracy, and false-accept/reject rates on safety-critical steps; use measurements to drive engineering decisions.
  • Build and maintain golden datasets and a non-regression suite, and use them as the release gate as the command and procedure catalogs grow.
  • Lead the migration from zero-shot to fine-tuned on-device models in order to reduce latency, without reintroducing a per-customer retraining burden.
  • Collaborate with product, firmware, and cloud teams, and bring new capabilities online, including additional languages, device commands, and guided workflows.

Educational Background

  • Trained in AI, Machine Learning, Electrical/Computer Engineering, or a related field, with specialization in NLP, speech, or deep learning; or equivalent industry experience delivering production conversational AI systems.
  • Applied background in NLU, dialog systems, or on-device machine learning.

Required Skills & Expertise

LLM and retrieval foundation — the baseline for this role

  • 5+ years of experience in ML/AI, with a strong focus on NLP, LLMs, or conversational AI.
  • Strong applied experience with LLMs: prompting, structured output, tool and function calling, evaluation, and retrieval-augmented generation (RAG) – together with the judgment to recognize when a model should not be used at all.
  • Solid command of embeddings and semantic similarity (e.g., cosine similarity, centroid versus maximum-similarity strategies, threshold tuning, and out-of-domain detection).
  • Strong Python with the ability to write clean, tested, reviewable code. Fluent with pytest, and Git and has CI discipline.

Edge dialog systems engineering — what this role additionally requires

  • Strong experience building edge conversational systems, including multi-turn dialog/state management and efficient intent/NLU pipelines using local-first, cheap-to-expensive inference strategies.
  • Disambiguation and repair: resolving ambiguous intent and noisy spoken references, and asking a clarifying question or re-prompting rather than committing to a confident wrong answer.
  • Comfort placing deterministic guardrails around a probabilistic model, including safety floors, confirmation gating, and negation handling, and experience with state-machine or workflow engines covering branching, variable capture, and resumability.
  • Experience running models on constrained hardware such as NPUs, mobile, or embedded targets, under real latency, memory budgets. This includes ONNX and onnxruntime, model quantization, and cross-architecture packaging for aarch64.
  • Practical embedded development workflow: Linux, Docker, adb, systemd services, and the ability to diagnose problems from device logs.
  • The ability to take ownership of an existing, non-trivial codebase and keep it healthy.
  • A safety-first instinct, a misheard answer on a safety-critical step is treated as a hazard, not as a metric regression.
  • Measurement discipline: benchmarking under realistic device conditions, curated golden datasets used as regression gates, and attention to embedding collisions and centroid drift as the command catalog grows.
  • Excellent problem-solving skills and strong written and verbal communication, with the ability to collaborate across engineering, product, and domain experts.

Growth axis — desirable rather than required

  • SLM fine-tuning, including LoRA and QLoRA, instruction and format tuning, and distillation of a larger evaluator model into a smaller on-device model.
  • Latency and footprint optimization: INT8 and INT4 quantization, ONNX export and graph optimization, hardware-aware model selection, and profiling to reduce inference costs.
  • A pragmatic view of the boundary between configuration-driven adaptation and fine-tuning.
  • Building the data flywheel that supports this work, turning on-device session logs into evaluation sets and training data.

Bonus Qualifications

  • Speech recognition experience and comfort working downstream of noisy transcription.
  • Familiarity with LLM-as-a-judge evaluation.
  • Multilingual NLU; the system's data layer is already language-partitioned.
  • Industrial, field-service, or safety-critical product experience, for example in utilities, energy, or manufacturing.
  • GO familiarity, for integration with the on-device voice pipeline agent.
  • Exposure to MCP or agentic tooling.
  • Prior experience in a startup or a fast-paced team, building products from ground up.

Location & Type

  • Full-time, on-site or hybrid, based in Palo Alto, CA.

Similar jobs

BrightAI CorporationPalo Alto, California, United States

Est. 200,000 USD

Sr. AI Engineer – LLM, RAG BrightAI is a high-growth Physical AI company transforming how businesses interact with the physical world through intelligent automation. Our AI platform processes visual, spatial, and tempora…

Full-time
SonatusDublin, Ireland

Est. 95,000 EUR

At Sonatus, we’re driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can’t keep pace with consumer expectations shaped by the mobile industry—where features evolve…

Full-time
N-iXRemote

We're looking for an engineer with hands-on experience building and evaluating GenAI services - from RAG and agentic reasoning systems to production-grade LLM deployments. You'll work closely with Frontend and Backend te…

Full-timeRemote
GradialSeattle, Washington, United States

Est. 175,000 USD

Gradial is the marketing operations system of work that helps marketers and creatives move from idea to execution faster. Our platform orchestrates across martech stacks, workflows, and people to automate marketing execu…

Full-time
LTSRemote

Est. 140,000 USD

Location: United States – RemoteClearance: Ability to obtain and maintain a Public Trust LTS is seeking a highly skilled Senior Applied AI Engineer to focus on continuously improving the intelligence behind the platform.…

Full-timeRemote
AvePointSingapore, Singapore

We are looking for a highly skilled AI Engineer specializing in Large Language Models (LLMs) and Agentic AI. You will architect, build, and deploy production-grade LLM applications — from intelligent knowledge bases and…

Full-time
Staff AI Engineer2 months ago
WorkatoRemote

About Workato Workato delivers enterprise infrastructure for the agentic era, redefining iPaaS and helping enterprises unify data, applications, processes, and AI into a single, governed platform. A leader in Enterprise…

Full-timeRemote
BrightAI CorporationPalo Alto, California, United States

Est. 141,000 USD

AI Engineer, Time-Series Signal Processing BrightAI is a high-growth Physical AI company transforming how businesses interact with the physical world through intelligent automation. Our platform processes visual, spatial…

Full-time
WorkatoSofia, Bulgaria

About Workato Workato delivers enterprise infrastructure for the agentic era, redefining iPaaS and helping enterprises unify data, applications, processes, and AI into a single, governed platform. A leader in Enterprise…

Full-time
PhizenixHyderabad, India

We are looking for an LLM / Agentic Evaluation Rig Engineer to build the system that decides whether our AI output is good enough to ship. Because our commentary sits next to externally reported financials, we cannot rel…

Full-time
ArionkoderRemote

Est. 120,000 USD

Arionkoder is an AI consulting firm helping companies build, embed, and own AI systems that drive real business outcomes. We combine Product Development, Artificial Intelligence, and Team Augmentation to craft digital pr…

Full-timeRemote
Senior AI Engineer4 months ago
SonatusDublin, Ireland

At Sonatus, we’re driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can’t keep pace with consumer expectations shaped by the mobile industry—where features evolve…

Full-time
Observe.AIRedwood City, California, United States

Est. 139,000 USD

About Us Observe.AI is the AI Agents platform for customer experience, designed to help organizations deliver faster, smarter, and more efficient customer service at scale. The platform enables businesses to deploy speci…

Full-time
Observe.AIRedwood City, California, United States

Est. 139,000 USD

About Us Observe.AI is the AI Agents platform for customer experience, designed to help organizations deliver faster, smarter, and more efficient customer service at scale. The platform enables businesses to deploy speci…

Full-time
LTSRemote

Est. 170,000 USD

Location: United States – RemoteClearance: Ability to obtain and maintain a Public Trust LTS is seeking a Senior Agentic AI Software Engineer to build the intelligence behind the platform—the autonomous agents, orchestra…

Full-timeRemote
Five9Bengaluru, India

Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-…

Full-time
GradialSeattle, Washington, United States

Est. 205,000 USD

Gradial is the marketing operations system of work that helps marketers and creatives move from idea to execution faster. Our platform orchestrates across martech stacks, workflows, and people to automate marketing execu…

Full-time
GradialSeattle, Washington, United States

Est. 175,000 USD

Gradial is the marketing operations system of work that helps marketers and creatives move from idea to execution faster. Our platform orchestrates across martech stacks, workflows, and people to automate marketing execu…

Full-time
AirPittsburgh, Pennsylvania, United States

Est. 124,000 USD

Company Description Air is the leader in Enterprise Readiness. Our mission is to establish readiness as a real-time condition that is continuously achieved. Today, a dangerous Readiness Gap exists between what the front…

Full-time
XPENGSanta Clara, California, United States

Est. 328,650 USD

XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and…

Full-time
WorkatoSan Francisco, California, United States

Est. 144,000 USD

About Workato Workato delivers enterprise infrastructure for the agentic era, redefining iPaaS and helping enterprises unify data, applications, processes, and AI into a single, governed platform. A leader in Enterprise…

Full-time
LiberateSan Francisco, California, United States

Est. 175,000 USD

About Us: Liberate builds AI agents to automate manual tasks for the $2.7T insurance industry. We started with voice — the hardest and most valuable channel in insurance — and are now expanding into full workflow automat…

Full-time
SonatusSunnyvale, California, United States

Est. 263,500 USD

At Sonatus, we’re driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can’t keep pace with consumer expectations shaped by the mobile industry—where features evolve…

Full-time
Hyphen Connect LimitedSan Francisco, California, United States

Est. 190,000 USD

We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…

Full-time
SonatusSunnyvale, California, United States

Est. 234,750 USD

At Sonatus, we’re driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can’t keep pace with consumer expectations shaped by the mobile industry—where features evolve…

Full-time
Hyphen Connect LimitedSeattle, Washington, United States

Est. 165,000 USD

We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…

Full-time
BackbaseVietnam

The Job in short The Data Enablement team is here to enable every team in the organisation with their data needs. Our job starts the moment that data enters our platform and ends when it reaches whoever needs it. We are…

Full-time
LiberateBoston, Massachusetts, United States

Est. 175,000 USD

About Us: Liberate builds AI agents to automate manual tasks for the $2.7T insurance industry. We started with voice — the hardest and most valuable channel in insurance — and are now expanding into full workflow automat…

Full-time
LTSRemote

Est. 140,000 USD

LTS is seeking an AI Platform and Harness Engineer to develop and maintain the infrastructure, tooling, and evaluation frameworks that power enterprise AI solutions. This role is responsible for building the AI platform…

Full-timeRemote
LTSRemote

Est. 144,000 USD

Location: United States – RemoteClearance: Ability to obtain and maintain a Public Trust LTS is seeking a highly skilled Agentic AI Security Engineer to ensure our AI systems are secure, trustworthy, resilient, and gover…

Full-timeRemote