Skip to content
FindAUseCase.com
  • Use cases
  • Technologies
  • Categories
  • Governance
  • Insights
  • About
Back to Directory

Amazon Elastic Kubernetes Service

1 use case using this technology

LogisticsLarge Language Models

Delhivery achieves 160 ms latency for high-precision geocoding using Amazon EKS

Delhivery

Delhivery, a logistics provider in India, implemented a fine-tuned open-source Llama 3.2 1B large language model on Amazon EKS to support high-volume geocoding of pickup and drop-off addresses. The system processes up to 8,000 requests per minute at 160 milliseconds latency using NVIDIA A10G GPU-backed G5 Xlarge instances and the vLLM framework. Delhivery cut model-serving costs by approximately 80 percent and accelerated prototyping cycles from two days to under six hours, working with the AWS Prototyping and Cloud Engineering (PACE) team.

FindAUseCase.com

The evidence library for private and European enterprise AI. Every entry sourced, scored and checked against a primary source.

Register

  • Browse use cases
  • Technology directory
  • Categories
  • Vendor directory
  • Partner directory
  • Governance register
  • Insights
  • FAQ

Company

  • About
  • Contact
  • Privacy policy
  • Terms of use
  • Cookie policy

Connect

TwitterLinkedInGitHubEmail

© 2026 FindAUseCase.com. All rights reserved.

Scores recomputed nightly