About this role
Backend Solution Engineer Location: Bengaluru (onsite) Experience: 4+ years
About Eka.care Eka.care builds healthcare technology for India, used by doctors and patients. Our AI Engineering team builds the systems behind EkaScribe , our AI medical scribe, which turns a doctor–patient conversation into a structured clinical document in real time. We also build the agentic systems that are starting to handle real clinical and operational workflows. These systems run in live clinics/ enterprise websites and government facilities. When they're slow or wrong, a user notices right away. We need engineers who treat reliability, scalability and correctness as a core essence.
The role We're looking for a strong backend and systems engineer who has hands-on exposure with AI. Engineering comes first and AI second. You'll design, build and own production systems that combine core engineering, streaming audio, LLMs and agents, and you'll run them at scale. You'll own features from design doc to production to on-call, and you'll ship fast without breaking things.
What you'll do
- Architect and build on EkaScribe's pipeline: streaming audio intake, ASR, transcript processing, LLM-based document generation, and the doctor's edit-and-publish loop. Low latency and high availability are requirements.
- Own reliability end to end: SLOs, fallbacks across model providers, retries, idempotency, backpressure, graceful degradation, incident response and postmortems.
- Build evals and observability for LLM systems: offline eval suites, online quality signals, hallucination and omission tracking, and LLM tracing. Every prompt or model change should ship with evidence that it's better.
- Build agentic systems: tool-using agents, multi-step orchestration, memory and retrieval, and guardrails that keep agents grounded in what was actually said and recorded.
- Design for scale and optimised solutions: queue-based async processing, caching, batching, token and GPU cost budgets, and multi-tenant isolation.
- Ship at a high pace: short iteration cycles, small safe deploys, feature flags, and clear written design docs.
- Work closely with product, clinical and mobile teams, and turn unclear clinical problems into well-scoped systems.
Must-have
- 4+ years building and running production backend systems , with a track record of owning services others depend on.
- Strong system design fundamentals: distributed systems, concurrency, consistency trade-offs, API design and data modelling.
- Strong exposure and adherence to system reliability principles
- High-speed delivery capability/ startup environment without compromising on reliability
- Hands-on with Python, Fast API, MCPs, Design Principles, AWS and Kubernetes : deploying, autoscaling and debugging services in production.
- Solid experience with Redis and Postgres : event-driven design, caching, and schema and query design at scale.
- Hands-on LLM application and agent experience in production: prompting, tool calling, structured outputs, RAG, and orchestration (LangGraph, CrewAI or custom).
- Experience building evals and observability for LLM systems: test sets, automated and LLM-as-judge scoring, tracing, and regression gates.
- A strong ownership mindset . You own it in production, not just on merge.
Nice-to-have
- Working knowledge of speech/ASR pipelines : streaming audio, ASR model trade-offs, diarization, and handling noisy, multilingual audio.
- Other languages / technology used for high-throughput services
- Vector search (pgvector or similar) and memory systems for agents
- Healthcare or regulated-data experience (ABDM, HIPAA, PII handling)
- Open-source contributions or public writing
Full-Time Employee Benefits: • Insurance Benefits - Medical Insurance, Accidental Insurance • Parental Support - Maternity Benefit, Paternity Benefit Program • Retirement Benefits - Employee PF Contribution, Gratuity, NPS, Leave Encashment • Other Benefits - Salary Advance Policy