keyLogo

DR

Hi, I'm Deep Rathi - AI Engineer @ Key | Ex‑ML Intern @ Remasto | Low‑Latency ML & Trading Systems | Generative AI | Real‑Time Inference & Optimization

I thrive on building and optimizing low-latency AI systems and pushing the boundaries of generative AI. My passion lies in translating complex AI research into production-ready solutions that drive real-time performance and innovation.

Key
Ahmedabad, Gujarat, India
AI insights

At a glance

Curated signals on strengths, focus areas, and how they can help.

Led machine learning initiatives as an intern at Remasto, optimizing real-time inference.

Currently building advanced generative AI and low-latency systems as an AI Engineer at Key.

Can help architect and deploy high-speed, low-latency AI solutions for critical applications.

🚀 Career trajectory

Emerging AI Leader

Building a strong foundation in AI engineering with a focus on performance and cutting-edge technologies.

AI Engineering

Low-Latency Systems

Generative AI

Applied AI Researcher

Gained practical experience through internships in Machine Learning and Data Science, applying AI across diverse domains.

Machine Learning

Data Science

Research

💪🏻 Superpowers

Low-Latency AI Systems Architect

Masters the art of building AI solutions where speed is paramount.

Expertise in optimizing real-time inference for applications like trading systems.

Proficient in multithreading and C++ for high-performance computing.

Drives efficiency in ML model deployment and execution.

Generative AI Innovator

Pioneering the development and application of cutting-edge generative AI.

Currently developing advanced generative AI models.

Leverages Large Language Models (LLMs) and frameworks like LangChain.

Translates AI research into practical, production-ready systems.

Production-Grade ML Deployment

Bridges the gap between AI research and real-world application.

Experience in deploying ML models into production environments.

Skilled in AWS for scalable AI infrastructure.

Focuses on turning AI concepts into tangible, high-impact systems.

I'm excited about

Developing next-generation generative AI models and optimizing their performance.

Building scalable, low-latency AI systems for critical applications.

Pushing the boundaries of real-time inference and AI optimization.

I can help with

Implementing high-speed, low-latency AI solutions for financial trading or real-time data processing.

Architecting and deploying generative AI models for innovative applications.

Optimizing existing ML pipelines for improved performance and efficiency.

I would love your help on

Connecting with innovators in the generative AI space.

Opportunities to tackle complex challenges in low-latency AI and distributed systems.

Collaborations on projects that require rapid AI deployment and real-time inference.

NotificationYou are testing the DEV environment.Create Issue