By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
AIModelKitAIModelKitAIModelKit
  • Home
  • News
    NewsShow More
    SpaceXAI’s Grok Tool Uploading Users’ Entire Codebase to Cloud Storage: What You Need to Know
    SpaceXAI’s Grok Tool Uploading Users’ Entire Codebase to Cloud Storage: What You Need to Know
    4 Min Read
    New York Leads the Way: First State to Enforce One-Year Moratorium on New AI Data Centers
    New York Leads the Way: First State to Enforce One-Year Moratorium on New AI Data Centers
    4 Min Read
    AI Replacing New York Nurses: Why Patients Should be Concerned About Quality of Care
    AI Replacing New York Nurses: Why Patients Should be Concerned About Quality of Care
    5 Min Read
    Navigating AI Agent Crawlers and Cloudflare’s New Rules: A Comprehensive Guide
    Navigating AI Agent Crawlers and Cloudflare’s New Rules: A Comprehensive Guide
    5 Min Read
    How Apple’s Self-Driving Car Program Paved the Way for Advanced AI Chip Technology
    How Apple’s Self-Driving Car Program Paved the Way for Advanced AI Chip Technology
    4 Min Read
  • Open-Source Models
    Open-Source ModelsShow More
    GlucoFM: Advanced Foundation Model for Continuous Glucose Monitoring Insights
    GlucoFM: Advanced Foundation Model for Continuous Glucose Monitoring Insights
    5 Min Read
    AgentHands: Creating Interactive Hand Gestures for Enhanced Conversations with Spatially Grounded Agents in XR
    AgentHands: Creating Interactive Hand Gestures for Enhanced Conversations with Spatially Grounded Agents in XR
    5 Min Read
    Exploring How Mobility Enhances Language Models’ Understanding of Location
    Exploring How Mobility Enhances Language Models’ Understanding of Location
    5 Min Read
    Optimize Candidate Biomarkers with Our AI Tool for Wearable Sensor Data Analysis
    Optimize Candidate Biomarkers with Our AI Tool for Wearable Sensor Data Analysis
    4 Min Read
    Beyond BMI: Assessing Cardiometabolic Risk Using Smartphone Images
    Beyond BMI: Assessing Cardiometabolic Risk Using Smartphone Images
    5 Min Read
  • Guides
    GuidesShow More
    Your Comprehensive Guide to Practical Constraint Decoding: Basics and Applications
    Your Comprehensive Guide to Practical Constraint Decoding: Basics and Applications
    6 Min Read
    KDnuggets Weekly Data Science News Roundup: Highlights from July 20, 2026
    KDnuggets Weekly Data Science News Roundup: Highlights from July 20, 2026
    4 Min Read
    Unlock Your AI Potential with Kaggle and Google’s Free 5-Day Agentic AI Course
    Unlock Your AI Potential with Kaggle and Google’s Free 5-Day Agentic AI Course
    6 Min Read
    Top 5 High-Performance MCP Servers for Optimal Agentic Development
    Top 5 High-Performance MCP Servers for Optimal Agentic Development
    6 Min Read
    Top 5 Free Resources for Understanding Agentic AI: Unlock Your Knowledge
    Top 5 Free Resources for Understanding Agentic AI: Unlock Your Knowledge
    6 Min Read
  • Tools
    ToolsShow More
    Unlock Agentic Coding: Experimenting with Qwen 3.8-Flash-Next on NVIDIA GB300 NVL72
    Unlock Agentic Coding: Experimenting with Qwen 3.8-Flash-Next on NVIDIA GB300 NVL72
    6 Min Read
    Unlock Agentic Coding: Experimenting with Qwen 3.8 Flash-Next 176B Model on NVIDIA GB300 NVL72
    Unlock Agentic Coding: Experimenting with Qwen 3.8 Flash-Next 176B Model on NVIDIA GB300 NVL72
    5 Min Read
    Optimizing LFM2.5 Q4_0 Checkpoints through Quantization-Aware Distillation Techniques
    Optimizing LFM2.5 Q4_0 Checkpoints through Quantization-Aware Distillation Techniques
    4 Min Read
    Deploy Qwen 3.8-2.4T-A95B: A Configurable 2.4T Parameter Model on NVIDIA GB300 NVL72 for Enhanced Reasoning
    Deploy Qwen 3.8-2.4T-A95B: A Configurable 2.4T Parameter Model on NVIDIA GB300 NVL72 for Enhanced Reasoning
    6 Min Read
    Optimize Your AI Models with Baseten on Hugging Face Inference Providers 🔥
    Optimize Your AI Models with Baseten on Hugging Face Inference Providers 🔥
    5 Min Read
  • Events
    EventsShow More
    Exploring the Future of EdTech: Highlights from the ‘Best of ISTE’ Virtual Playground
    Exploring the Future of EdTech: Highlights from the ‘Best of ISTE’ Virtual Playground
    4 Min Read
    Empowering Veteran Students: Effective Teaching Strategies in Technology and Learning
    Empowering Veteran Students: Effective Teaching Strategies in Technology and Learning
    4 Min Read
    NVIDIA Partners with NSF to Enhance AI Research and Education Through State and Regional AI Hubs Across the US
    NVIDIA Partners with NSF to Enhance AI Research and Education Through State and Regional AI Hubs Across the US
    5 Min Read
    South Korea Unveils AI Future at AI Summit with NVIDIA and Strategic Partners
    South Korea Unveils AI Future at AI Summit with NVIDIA and Strategic Partners
    5 Min Read
    NVIDIA Launches First Open-Source GPU-Accelerated Framework for Medical Physics Simulations
    NVIDIA Launches First Open-Source GPU-Accelerated Framework for Medical Physics Simulations
    5 Min Read
  • Ethics
    EthicsShow More
    AI Giants Warn: Impending Cybersecurity Crisis Looms in Just Months
    AI Giants Warn: Impending Cybersecurity Crisis Looms in Just Months
    6 Min Read
    Assessing the Environmental Impact of Data Centres: Are We Finally Acknowledging the Consequences?
    Assessing the Environmental Impact of Data Centres: Are We Finally Acknowledging the Consequences?
    5 Min Read
    Survey Reveals Surprising Impact of AI on Job Losses: Insights from Workers
    Survey Reveals Surprising Impact of AI on Job Losses: Insights from Workers
    6 Min Read
    Understanding DAO-to-DAO Voting: On-Chain and Off-Chain Mechanisms Explored
    Understanding DAO-to-DAO Voting: On-Chain and Off-Chain Mechanisms Explored
    5 Min Read
    Taiwan Prosecutes Nine Individuals for Smuggling Advanced AI Servers to China: A Tech Industry Update
    Taiwan Prosecutes Nine Individuals for Smuggling Advanced AI Servers to China: A Tech Industry Update
    4 Min Read
  • Comparisons
    ComparisonsShow More
    InternBootcamp: Enhancing LLM Reasoning Through Verifiable Task Scaling Techniques
    InternBootcamp: Enhancing LLM Reasoning Through Verifiable Task Scaling Techniques
    4 Min Read
    Enhancing Anomaly Detection in Collider Experiments through Contrastive Learning for Better Interpretability
    Enhancing Anomaly Detection in Collider Experiments through Contrastive Learning for Better Interpretability
    6 Min Read
    Exploring the Impact of Quantization on Self-Explanations in Large Language Models: Can LLMs Explain Themselves?
    Exploring the Impact of Quantization on Self-Explanations in Large Language Models: Can LLMs Explain Themselves?
    5 Min Read
    CytoNet: A Foundation Model for Understanding the Human Cerebral Cortex at Cellular Resolution
    CytoNet: A Foundation Model for Understanding the Human Cerebral Cortex at Cellular Resolution
    5 Min Read
    Optimizing Nonconvex-Nonconcave Min-Max Problems with a Limited Maximization Domain: Insights from [2110.03950]
    Optimizing Nonconvex-Nonconcave Min-Max Problems with a Limited Maximization Domain: Insights from [2110.03950]
    5 Min Read
Search
  • Privacy Policy
  • Terms of Service
  • Contact Us
  • FAQ / Help Center
  • Advertise With Us
  • Latest News
  • Model Comparisons
  • Tutorials & Guides
  • Open-Source Tools
  • Community Events
© 2025 AI Model Kit. All Rights Reserved.
Reading: Swiggy Unveils Hermes V3: Transforming Text-to-SQL Into Conversational AI Solutions
Share
Notification Show More
Font ResizerAa
AIModelKitAIModelKit
Font ResizerAa
  • 🏠
  • 🚀
  • 📰
  • 💡
  • 📚
  • ⭐
Search
  • Home
  • News
  • Models
  • Guides
  • Tools
  • Ethics
  • Events
  • Comparisons
Follow US
  • Latest News
  • Model Comparisons
  • Tutorials & Guides
  • Open-Source Tools
  • Community Events
© 2025 AI Model Kit. All Rights Reserved.
AIModelKit > Comparisons > Swiggy Unveils Hermes V3: Transforming Text-to-SQL Into Conversational AI Solutions
Comparisons

Swiggy Unveils Hermes V3: Transforming Text-to-SQL Into Conversational AI Solutions

aimodelkit
Last updated: January 2, 2026 5:00 pm
aimodelkit
Share
SHARE

Introducing Hermes V3: Swiggy’s GenAI-Powered Text-to-SQL Assistant

Swiggy has made significant strides in enhancing data accessibility for its employees by launching Hermes V3, a GenAI-powered text-to-SQL assistant. This innovative tool allows users to query complex datasets effortlessly using plain English, marking a remarkable evolution from its earlier versions. Available directly within Slack, Hermes V3 combines cutting-edge technologies such as vector retrieval, session memory, agentic orchestration, and an explanation layer to transform natural language inputs into accurate SQL queries.

From Simple Queries to Advanced Interactions

Initially, Hermes served as a lightweight interface designed for straightforward questions and corresponding SQL queries executed against Swiggy’s internal data repositories. However, the early versions of this assistant faced considerable limitations. They struggled with derived metrics and lacked the conversational context necessary for creating consistent results across similar prompts. Moreover, users had no clear method for validating the generated SQL.

The Rebuild: Addressing Key Challenges

To tackle these challenges, Swiggy’s engineering team embarked on a revamp of Hermes using advanced machine learning techniques. By leveraging few-shot learning, metadata retrieval, and structured workflows centered on large language models, they effectively rectified previous shortcomings. This allowed for a significantly enhanced interaction experience and overall functionality.

Previous Hermes architecture

Previous Hermes overall architecture (Source: Swiggy Tech Blog)

Unleashing Vector-Based Retrieval for Enhanced Accuracy

With the launch of Hermes V3, a vector-based prompt retrieval system has been introduced. This system is built on historical SQL executed in Snowflake, addressing the previous challenges posed by missing descriptive metadata in production queries. By harnessing large-context language models, Hermes V3 can now convert SQL queries into natural-language explanations, allowing the system to reconstruct the query intent effectively.

As stated by Meghana Negi and Rutvik Reddy, Engineers at Swiggy:

“Hermes now taps into a curated database of previously executed queries and their prompts, uses vector similarity for retrieval, and remembers conversational context, improving SQL generation accuracy from 54% to 93% while enabling natural, multi-turn interactions.”

Multi-Turn Queries and Enhanced Contextual Memory

One of the standout features of Hermes V3 is its ability to maintain conversational memory. This capability allows users to engage in multi-turn interactions without needing to repeatedly establish context. The system adeptly tracks session states, translating simple metrics into compound requests. An orchestrator agent implements a ReAct-style reasoning loop, efficiently breaking down complex questions into manageable tasks—ranging from intent parsing to SQL generation.

Hermes V3 workflow

Hermes V3 workflow (Source: Swiggy Tech Blog)

The Explanation Layer: Enhancing Trust with Transparency

Another significant upgrade in Hermes V3 is the addition of an explanation layer. This feature surfaces the assumptions behind generated SQL queries and assigns confidence scores, empowering non-technical stakeholders to grasp the mechanics of how queries are formed. This newfound transparency fosters an environment of trust, ensuring users are confident in the machine-generated insights.

Securing Data Access with Robust Protocols

Security and compliance are paramount for Swiggy, and Hermes V3 is no exception. The system is seamlessly integrated with Swiggy’s stringent security measures, including role-based access control, single sign-on capabilities, and audit logs. This ensures that sensitive data access adheres strictly to internal governance policies. Additionally, hybrid metadata retrieval strategies are employed to efficiently fetch relevant schema, tables, and column details without exceeding token usage limits.

Architechting with Open-Source and Cloud-Native Technologies

The architecture of Hermes intertwines multiple open-source and cloud-native technologies. Vector databases and embedding models facilitate robust retrieval functions, while workflow orchestration makes extensive use of tools like LangChain. Observability frameworks have been layered on, enhancing provenance and monitoring features. Moreover, tools such as Snowflake for analytics and PostgreSQL for database management form the backbone of the ecosystem that supports Hermes’s functionality.

With these groundbreaking enhancements, Hermes V3 has solidified itself as an indispensable tool within Swiggy, allowing employees to derive valuable insights from complex datasets with unprecedented ease and accuracy.

Inspired by: Source

Contents
  • From Simple Queries to Advanced Interactions
  • The Rebuild: Addressing Key Challenges
  • Unleashing Vector-Based Retrieval for Enhanced Accuracy
  • Multi-Turn Queries and Enhanced Contextual Memory
  • The Explanation Layer: Enhancing Trust with Transparency
  • Securing Data Access with Robust Protocols
  • Architechting with Open-Source and Cloud-Native Technologies
Cloudflare Launches AI-Powered Experimental Alternative to Next.js
Scaling Discord’s ML Platform: From Single-GPU Workflows to a Shared Ray Cluster Setup
Unlocking LLM Mathematical Reasoning: Analyzing Frequency-Domain Fingerprints
Enhanced Segmentation of Cellular-Potts Agent-Based Models Using U-Net Neural Network: A Surrogate Modeling Approach
Google Unveils Gemini CLI: An Open-Source Terminal AI Agent Designed for Developers

Sign Up For Daily Newsletter

Get AI news first! Join our newsletter for fresh updates on open-source models.

By signing up, you agree to our Terms of Use and acknowledge the data practices in our Privacy Policy. You may unsubscribe at any time.
Share This Article
Facebook Copy Link Print
Previous Article Amazon S3 Vectors Achieves General Availability: Unveiling a ‘Storage-First’ Architecture for Retrieval-Augmented Generation (RAG) Amazon S3 Vectors Achieves General Availability: Unveiling a ‘Storage-First’ Architecture for Retrieval-Augmented Generation (RAG)
Next Article January 2026 EdTech Show & Tell: Innovations in Educational Technology January 2026 EdTech Show & Tell: Innovations in Educational Technology

Stay Connected

XFollow
PinterestPin
TelegramFollow
LinkedInFollow

							banner							
							banner
Explore Top AI Tools Instantly
Discover, compare, and choose the best AI tools in one place. Easy search, real-time updates, and expert-picked solutions.
Browse AI Tools

Latest News

AI Giants Warn: Impending Cybersecurity Crisis Looms in Just Months
AI Giants Warn: Impending Cybersecurity Crisis Looms in Just Months
Ethics
Assessing the Environmental Impact of Data Centres: Are We Finally Acknowledging the Consequences?
Assessing the Environmental Impact of Data Centres: Are We Finally Acknowledging the Consequences?
Ethics
Exploring the Future of EdTech: Highlights from the ‘Best of ISTE’ Virtual Playground
Exploring the Future of EdTech: Highlights from the ‘Best of ISTE’ Virtual Playground
Events
Survey Reveals Surprising Impact of AI on Job Losses: Insights from Workers
Survey Reveals Surprising Impact of AI on Job Losses: Insights from Workers
Ethics
//

Leading global tech insights for 20M+ innovators

Quick Link

  • Latest News
  • Model Comparisons
  • Tutorials & Guides
  • Open-Source Tools
  • Community Events

Support

  • Privacy Policy
  • Terms of Service
  • Contact Us
  • FAQ / Help Center
  • Advertise With Us

Sign Up for Our Newsletter

Get AI news first! Join our newsletter for fresh updates on open-source models.

AIModelKitAIModelKit
Follow US
© 2025 AI Model Kit. All Rights Reserved.
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?