This listing is optimized for AI-native talent: clear skills and tools so candidates know if they fit before they apply—and so you spend less time screening mismatches.
Quick answers for candidates and hiring teams reviewing this listing.
You optimize answer engines that combine live retrieval, citations, and LLMs—quality and latency at query scale matter more than offline benchmarks alone.
Hybrid SF is typical. Confirm schedule and US work authorization with the team.
Link writeups showing retrieval improvements, citation accuracy fixes, or cost/latency wins on search-like products.
Suggested
AI talent you may also like
Profiles ranked by overlap with this role's skills and tools—handy if you're hiring for a team or comparing backup candidates.
No close matches in the directory yet—try Find talent with your own filters.
Similar open roles
Other listings that share skills or tools with this one—useful if you want comparable stacks or backup options.
Cohere builds enterprise-grade language models and APIs used globally from our Toronto headquarters.Join as a Senior LLM Engineer to ship RAG, embeddings, and fine-tuning workflows for regulated customers.What you will doShip production LLM and retrieval systems with offline evals, monitoring, and…
Perks · GPU access, equity, Toronto waterfront office.
PythonLangChainOpenAI APIPyTorch
ToolsVector DBsDocker
View role
Listed Jul 27, 2026
Updated Aug 3, 2026
Hugging Face héberge l'écosystème open-source ML le plus utilisé au monde — équipe produit à Paris.Senior RAG Engineer: retrieval, Spaces, et endpoints inference avec latence et qualité mesurées.MissionsAméliorer les pipelines RAG (chunking, embeddings, hybrid search).Publier des patterns…
BI SOLUTIONS intervient sur des plateformes RAG en IDF avec exigences de performance et gouvernance des données.Mission orientée Python + retrieval, avec forte autonomie et ownership production.Ce que vous allez faireLivrer en production des pipelines RAG avec contraintes de coût et latence avec…
Perks · Mission remote IDF, forte autonomie technique.
PythonLangChainSQL
ToolsVector DBsDocker
View role
Listed Jul 12, 2026Updated Aug 3, 2026
Glean connects enterprise knowledge with LLM search and agents used by Fortune 500 teams.LLM Engineer — enterprise RAG (Palo Alto / hybrid): connectors, ACL-aware retrieval, and quality evals.What you will doDeliver production-grade enterprise RAG with measurable quality, latency, and cost…
Perks · Equity, hybrid, comprehensive medical.
PythonLangChainVector DBs
ToolsDocker
View role
Listed Jul 3, 2026Updated Aug 3, 2026
WOP360 is a global news platform covering 195 countries — politics, economy, technology, sport, and breaking briefings driven by real-world search trends.Senior AI Developer (Boston hybrid): build NLP and LLM systems that help editors detect trending topics, draft country-desk briefings, and…
Perks · Hybrid Boston, health coverage, conference budget, global newsroom access.