{"schemaVersion":"jobsearcher.job.v1","id":"adeefaa6d9954d8a487eb1b6","url":"https://jobsearcher.com/jobs/adeefaa6d9954d8a487eb1b6","canonicalUrl":"https://jobsearcher.com/jobs/adeefaa6d9954d8a487eb1b6","title":"AI/ML Engineer","description":"ML Engineer — Early Role at Pre-Seed StartupCyrano is an early-stage, pre-seed, revenue-generating AI startup in SF building groundbreaking technology for consumer and business use. We are looking for early technical contributors who live and breathe AI, tools, frameworks, and models.You'll take ambiguous problems from idea to production across language-model features, agent workflows, retrieval, memory, and evaluation, then prove the thing works and make it better.Qualifications We want engineers obsessed with building, evaluating, and iterating to deliver something new. These are required.Advanced proficiency in Python, with a strong emphasis on async frameworks (asyncio, FastAPI, AIOHTTP) built to handle concurrent, streaming WebSocket connections, plus working proficiency in at least one of C++, Rust, or Go.Real CS foundations: graph theory, relational theory, algorithmic complexity, and big-O analysis you can actually reason with, not just recite.Experience building and shipping LLM-powered features: RAG pipelines (chunking, embeddings, vector databases, reranking), tool use, and agentic workflows.Direct experience building, orchestrating, and evaluation-testing autonomous agents using frameworks like LangChain, AutoGen, CrewAI, LlamaIndex, or custom DAG-based workflows.Practical deep learning: transformer architectures and a framework such as PyTorch, JAX, or TensorFlow.Fine-tuning and post-training (supervised fine-tuning, LoRA/QLoRA, preference optimization), including the judgment to know when not to fine-tune.Experience implementing short-term, long-term, and episodic memory architectures (e.g., Zep, Mem0, semantic search context windows) for persistent conversational state.Production infrastructure: containers, Kubernetes, a major cloud platform, automated build/test/deploy pipelines, and model observability.Evaluation rigor: you build eval datasets, measure grounding and hallucination failures, catch regressions, and reason about quality, latency, throughput, and cost together.Nice to Have Additional strengths. They don't replace the list above.Experience optimizing and serving LLMs and speech models using frameworks like vLLM, TensorRT-LLM, Ollama, or Triton Inference Server.Real-time voice or speech applications and concurrent streaming systems.Distributed training, open-source contributions, publications, or a portfolio of shipped AI products.What We Look ForEnd-to-end ownership: comfort taking ambiguous problems from idea to production.Clear written and verbal communication with technical and non-technical stakeholders. You use Git, document trade-offs, and land reliable deliverables.Fast learner who keeps pace with a rapidly evolving model landscape.We value demonstrated capability across these areas, not a list of tools alone. A degree in computer science, AI/ML, or a related field may be useful but is not required.Work Arrangement & Compensation Full-time, in person. San Francisco Bay Area residence required. Base salary $120,000–$200,000 USD, plus early equity for the right candidate.Apply: https://docs.google.com/forms/d/e/1FAIpQLSdrsnW2CpvYDBRZnS91cL4nXgnvi1mwQSYqF6IhvGPImaLKYw/viewform","company":"Cyrano","rawCompany":"cyrano","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-28T09:47:13.968Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"AI/ML Engineer","description":"ML Engineer — Early Role at Pre-Seed StartupCyrano is an early-stage, pre-seed, revenue-generating AI startup in SF building groundbreaking technology for consumer and business use. We are looking for early technical contributors who live and breathe AI, tools, frameworks, and models.You'll take ambiguous problems from idea to production across language-model features, agent workflows, retrieval, memory, and evaluation, then prove the thing works and make it better.Qualifications We want engineers obsessed with building, evaluating, and iterating to deliver something new. These are required.Advanced proficiency in Python, with a strong emphasis on async frameworks (asyncio, FastAPI, AIOHTTP) built to handle concurrent, streaming WebSocket connections, plus working proficiency in at least one of C++, Rust, or Go.Real CS foundations: graph theory, relational theory, algorithmic complexity, and big-O analysis you can actually reason with, not just recite.Experience building and shipping LLM-powered features: RAG pipelines (chunking, embeddings, vector databases, reranking), tool use, and agentic workflows.Direct experience building, orchestrating, and evaluation-testing autonomous agents using frameworks like LangChain, AutoGen, CrewAI, LlamaIndex, or custom DAG-based workflows.Practical deep learning: transformer architectures and a framework such as PyTorch, JAX, or TensorFlow.Fine-tuning and post-training (supervised fine-tuning, LoRA/QLoRA, preference optimization), including the judgment to know when not to fine-tune.Experience implementing short-term, long-term, and episodic memory architectures (e.g., Zep, Mem0, semantic search context windows) for persistent conversational state.Production infrastructure: containers, Kubernetes, a major cloud platform, automated build/test/deploy pipelines, and model observability.Evaluation rigor: you build eval datasets, measure grounding and hallucination failures, catch regressions, and reason about quality, latency, throughput, and cost together.Nice to Have Additional strengths. They don't replace the list above.Experience optimizing and serving LLMs and speech models using frameworks like vLLM, TensorRT-LLM, Ollama, or Triton Inference Server.Real-time voice or speech applications and concurrent streaming systems.Distributed training, open-source contributions, publications, or a portfolio of shipped AI products.What We Look ForEnd-to-end ownership: comfort taking ambiguous problems from idea to production.Clear written and verbal communication with technical and non-technical stakeholders. You use Git, document trade-offs, and land reliable deliverables.Fast learner who keeps pace with a rapidly evolving model landscape.We value demonstrated capability across these areas, not a list of tools alone. A degree in computer science, AI/ML, or a related field may be useful but is not required.Work Arrangement & Compensation Full-time, in person. San Francisco Bay Area residence required. Base salary $120,000–$200,000 USD, plus early equity for the right candidate.Apply: https://docs.google.com/forms/d/e/1FAIpQLSdrsnW2CpvYDBRZnS91cL4nXgnvi1mwQSYqF6IhvGPImaLKYw/viewform","datePosted":"2026-09-28T09:47:13.968Z","dateModified":"2026-09-28T09:47:13.968Z","hiringOrganization":{"@type":"Organization","name":"Cyrano","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"adeefaa6d9954d8a487eb1b6"},"url":"https://jobsearcher.com/jobs/adeefaa6d9954d8a487eb1b6"}}