876 to 900 of 1,256 Permanent Low Latency Jobs

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
El Paso, Texas, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Providence, Rhode Island, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Fort Myers, Florida, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Grand Rapids, Michigan, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Overland Park, Kansas, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Little Rock, Arkansas, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Las Vegas, Nevada, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Fort Worth, Texas, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Des Moines, Iowa, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Charleston, South Carolina, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Cedar Rapids, Iowa, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Kansas City, Missouri, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Rapid City, South Dakota, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Salt Lake City, Utah, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Sr. AI Engineer (Speech)

Hiring Organisation
Dialpad
Location
Sioux Falls, South Dakota, United States
Employment Type
Permanent
Salary
USD Annual
including decoding, endpointing, turn detection, streaming behavior, and the handoff between ASR, LLM, and TTS, with a focus on natural interactions and low latency. Research & Evaluation: Benchmark, prototype, fine-tune, distill, or otherwise adapt models and algorithms when doing so can create a meaningful advantage in voice-agent quality … latency, cost, or reliability. Production ML & Backend: Design and ship production-grade services and inference components for real-time speech, partnering with platform and backend engineers to make model improvements observable, scalable, and maintainable. Cross-Functional Leadership & Mentorship: Partner with speech, NLP, telephony, platform, product, and infrastructure engineers ...

Principal Platform Software Engineer, C++

Hiring Organisation
Evolv Technologies Inc
Location
Cambridge, Massachusetts, United States
Employment Type
Permanent
Salary
USD Annual
C++ software components that follow industry-standard design patterns, development methodologies, and deployment models Design and implement real-time systems with deterministic performance and low-latency requirements. Develop multithreaded applications using modern C++ concurrency primitives, thread synchronization, and lock-free programming. Integrate ML inference engines into fast-path … pipelines with minimal latency impact while optimizing control software for embedded systems and real-time hardware interfaces. Perform technical performance benchmarking and analyses to support engineering decisions. Write clean, well-documented, and testable code. Maintain quality throughout software development through peer code review, unit and functional testing. Leadership, Team ...

Automation-Driven Network Engineer (Low-Latency)

Location
Greater London, England, United Kingdom
Marex Group is seeking a strong network engineer to join the Network Team in a fast-paced fintech environment in Greater London. The role focuses on the end-to-end design, deployment and operation of ...

Rates eTrading Strategist: Low-Latency Quant Trader

Location
Greater London, England, United Kingdom
LGBT Great is looking for an innovative eTrading Strategist to join their team in Greater London. The ideal candidate will design and develop algorithmic trading strategies, contributing to the overall electronic trading infrastructure. The role ...

Low-Latency Quant Developer (Java/Rust) — Hybrid

Location
Greater London, England, United Kingdom
Citi is seeking a Quantitative Analyst/Developer to join our electronic execution team and drive the development of cash equity algorithmic trading platforms. You will design and optimize high-performance trading systems using Java ...

Senior Lead Software Engineer — Low-Latency Trading Platform

Location
Greater London, England, United Kingdom
慨正橡扯 is seeking a Senior Software Engineer to drive innovation within the Electronic Trading Services Platforms in Greater London. You will lead multiple complex projects, mentor teams, and align technical strategies with business goals. The ...

Principal Software Engineer, AI Platform Engineering

Hiring Organisation
Saviynt
Location
Milpitas, California, United States
Employment Type
Permanent
Salary
USD Annual
operate Pgvector (Cloud SQL) for POC and Qdrant on GKE for production-scale embedding storage; design index strategies (IVFFlat, HNSW) and manage ANN query latency SLAs RAG data pipeline: build embedding generation pipelines that chunk, encode, and upsert document embeddings into the vector store; own the data refresh cadence … staleness SLAs for retrieval context Service APIs: expose data platform services (feature serving, embedding upsert, schema validation) over HTTPS with mTLS and gRPC where low-latency streaming is required Synthetic data pipelines for dev/staging where real customer data is not permitted Data quality gates: Great Expectations ...

Principal Software Engineer, AI Platform Engineering

Hiring Organisation
Saviynt
Location
San Jose, California, United States
Employment Type
Permanent
Salary
USD Annual
operate Pgvector (Cloud SQL) for POC and Qdrant on GKE for production-scale embedding storage; design index strategies (IVFFlat, HNSW) and manage ANN query latency SLAs RAG data pipeline: build embedding generation pipelines that chunk, encode, and upsert document embeddings into the vector store; own the data refresh cadence … staleness SLAs for retrieval context Service APIs: expose data platform services (feature serving, embedding upsert, schema validation) over HTTPS with mTLS and gRPC where low-latency streaming is required Synthetic data pipelines for dev/staging where real customer data is not permitted Data quality gates: Great Expectations ...

Staff HPC Software Engineer

Hiring Organisation
San Diego Stealth Startup
Location
San Diego, California, United States
Employment Type
Permanent
Salary
USD Annual
every algorithm or infrastructure service. Their core responsibility is making the data and compute path reliable under real throughput, storage, network, and latency constraints. Early work In the first three to six months, this person should help: Establish reproducible hardware benchmarks for accelerated compute, CPU workloads, memory transfers, storage … CUDA, C, or C#. Comfortable with the normal engineering tools: profiling, tracing, debugging, testing, code review, builds, and CI. Can reason concretely about throughput, latency, buffering, memory, storage, network behavior, scheduling, contention, and failure recovery. Uses measurements to guide performance work: can identify a bottleneck, make a targeted change ...

Embedded Systems Engineer, Humanoid Robotics

Hiring Organisation
FieldAI
Location
Quincy, Massachusetts, United States
Employment Type
Permanent
Salary
USD Annual
closely with the ML team building the robot's software brain, ensuring the compute platform can run their perception and manipulation models with the latency and throughput they need. The system is designed to operate across a diversity of humanoid robot platforms, so your work will generalize across different … level to support real-time performance and reliable operation of the backpack's compute stack. Testing & Diagnostics: Conduct thermal profiling, power draw analysis, and latency measurement, and implement watchdogs and health checks for the compute stack. 2. Sensor & Actuator Drivers Perception & State Sensor Drivers: Adapt, integrate, and where needed ...

Senior Software Engineer, Mapping

Hiring Organisation
ALSO
Location
San Jose, California, United States
Employment Type
Permanent
Salary
USD Annual
Collaborate with firmware, mobile, and product teams to define the navigation system interface between bike hardware, the mobile app, and cloud services. Design robust, low-latency APIs and cloud systems that handle location, routing logic, vehicle-specific optimizations, and data synchronization. Enable critical edge cases like offline fallback … systems in a fast-moving startup or 0 1 product environment. Deep understanding of real-time or near real-time system design-especially in latency-sensitive use cases. Extensive experience with mapping and navigation platforms and integration (e.g., Google Maps APIs, Mapbox, HERE, TomTom) Excellent technical judgment ...