Bengaluru · Staff/Principal
Applicants who checked fit first are 3.1× more likely to hear back
Your score for this role already exists
ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.
No credit card · 1 tap with Google
LOW
JioStar is showing limited hiring activity lately. Expect slight delayed response.
First 72 hours
Window passed - posted 130d agoEarly applicants get seen before the pile builds.
Not a repost
The first time we've seen this listing - it hasn't been closed and reopened.
You almost certainly match several of these already. Unlock your skill map to see the matches, the gaps, and what to fix first.
You are a retrieval-oriented engineer with deep expertise in high-dimensional data, relational structures, and large-scale knowledge representation. You thrive on the challenge of bridging the gap between raw data and semantic understanding, building the backbone for next-generation AI and discovery systems. You are passionate about data topology, latent space optimization, and the performance tuning of complex query engines. You constantly strive to reduce "time-to-insight" and maximize the precision of information retrieval at scale.
The pace of our growth is incredible—if you want to tackle the foundational challenges of RAG (Retrieval-Augmented Generation), knowledge graphs, and semantic search at a global scale, join us!
Lead the design and development of hybrid retrieval architectures combining vector similarity search with structured graph traversals.
Architect scalable data pipelines for the ingestion, embedding, and indexing of massive, multi-modal datasets.
Innovate and prototype advanced retrieval techniques, including multi-stage re-ranking, graph-tooling for LLMs, and dynamic metadata filtering.
Design and implement schemas for complex knowledge graphs, ensuring high-performance relationship mapping and ontological integrity.
Build automated data validation and drift detection systems to monitor the quality of embeddings and the health of the vector space.
Drive technical implementation of "Memory" systems for AI agents, focusing on long-term persistence, observability, and sub-second latency.
Champion data organization standards, ensuring that disparate data sources are unified into a coherent, searchable knowledge base.
Collaborate with AI Research and Product teams to evaluate emerging database technologies (e.g., HNSW optimizations, GraphRAG) and integrate them into production.
7+ years of experience in data engineering or backend systems with a focus on high-performance data retrieval and storage.
BE/B.Tech in Computer Science, Mathematics, or equivalent. MS or PhD in a related field is a plus.
Expert proficiency in Python, Java, or Go, with a strong grasp of distributed system design patterns.
Deep understanding of Vector Databases, including indexing strategies (HNSW, IVFFlat, PQ) and distance metrics (Cosine, Euclidean, Dot Product). Experience with Pinecone, Milvus, Weaviate, or Qdrant.
Strong background in Graph Databases (Neo4j, AWS Neptune, or ArangoDB) and query languages like Cypher or Gremlin.
Experience with Data Modeling and organization, specifically in building semantic layers, ontologies, and taxonomies.
Hands-on experience with LLM orchestration frameworks (LangChain, LlamaIndex) and embedding models (OpenAI, HuggingFace, Cohere).
Proficiency in large-scale data processing using Spark, Flink, or Kafka for real-time indexing and ETL.
Understanding of Information Retrieval (IR) fundamentals, including BM25, TF-IDF, and reciprocal rank fusion.
Experience with cloud-native infrastructure (AWS/GCP/Azure) and container orchestration (Kubernetes).
Free · no signup
Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.
No spam. Just jobs and resources.
Why people use ASAI
Scored, not searched. Every role ranked against your actual profile.
Alerts as often as hourly. Reach new roles while the pile is still small.
Skill gaps, spelled out. See exactly which requirements you don't meet yet.
Verified jobs, only. Say no to ghost jobs. Your time deserves respect.
More Data Engineer roles in Bengaluru
See allKeep browsing