Skip to content
View ayushguptax's full-sized avatar
  • India
  • 11:47 (UTC +05:30)

Block or report ayushguptax

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
ayushguptax/README.md

Ayush Gupta

Staff Software Engineer | Distributed Systems & AI Architect

Building high-scale real-time platforms, distributed AI inference clusters, and resilient cloud architectures.

Portfolio


Overview

Staff Software Engineer and Founding Engineer with 5+ years of experience architecting high-throughput distributed platforms, real-time media systems, and production AI infrastructure. Driven by a "Systems First" engineering philosophy — prioritizing sub-10ms latencies, high availability, zero-trust observability, and rapid developer velocity.


Production Scale & Core Impact

  • Infrastructure at Scale: Architected zero-to-one platform infrastructure for Vooz Inc., scaling to 500,000+ Monthly Active Users (MAU) with zero downtime.
  • Distributed AI Inference: Migrated standalone ML models to an NVIDIA Triton Inference Server cluster featuring dynamic batching and GPU optimization, scaling throughput to 1,280+ RPS.
  • Edge Intelligence: Engineered zero-latency, client-side browser inference using ONNX Runtime (WebGPU / WASM) and Cache API for offline-capable AI features.
  • Real-Time Financial Settlement: Designed Solana microservices integrated with Helius Webhooks for atomic payment confirmations and real-time point balance hydration.
  • Resilient Media Recovery: Built automated WebRTC connection recovery mechanisms and state synchronization for fault-tolerant audio/video streaming.

Technical Architecture & Systems Stack

Core Languages

Go Python TypeScript C++

AI Infrastructure & Inference

NVIDIA Triton ONNX WebGPU PyTorch Qdrant Memgraph

Backend & Distributed Systems

Node.js FastAPI gRPC Redis PostgreSQL

Real-Time & Media Systems

WebRTC SignalR Socket.io Next.js 16

Cloud, DevOps & Observability

Kubernetes Azure AKS ArgoCD Docker OpenTelemetry AWS


Engineering Principles & Architecture

  1. Systems-First Resilience: Design for explicit failure modes, distributed tracing, and graceful degradation under extreme load.
  2. Edge ML Acceleration: Offload inference compute to browser-local WebGPU/WASM models to eliminate network round-trips.
  3. Immutable Infrastructure: Enforce GitOps continuous deployment, infrastructure as code, and declarative cluster management.
  4. Data-Driven Performance Tuning: Rely on OpenTelemetry profiling and metrics to resolve micro-bottlenecks before scaling horizontal compute.

Popular repositories Loading

  1. Emoji-Picker Emoji-Picker Public

    A simple emoji picker component for React & Next apps

    TypeScript 4

  2. react-spring-bottom-sheet react-spring-bottom-sheet Public

    ✨ Accessible, 🪄 Delightful, and 🤯 Performant, updated to support React 19.

    TypeScript 1

  3. ayushguptax ayushguptax Public

  4. school-emails-lookup school-emails-lookup Public

    JavaScript

  5. bitchat bitchat Public

    Forked from permissionlesstech/bitchat

    bluetooth mesh chat, IRC vibes

    Swift