Skip to content
View DhruvGarg111's full-sized avatar

Highlights

  • Pro

Block or report DhruvGarg111

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
DhruvGarg111/README.md

Dhruv Garg

Building practical AI systems, one focused iteration at a time.

Roles

πŸ”¬ Engineering Profile

Machine Learning Engineer focused on Computer Vision and Agentic AI, backed by scalable distributed infrastructure. My focus is translating cutting-edge ML research into optimized, production-grade systems.

  • 🎯 Vision Systems: Bypassing computational bottlenecks in 4K aerial detection via Explainable AI (LayerCAM).
  • πŸ€– Agentic AI: Engineering autonomous tool-using agents & local LLM orchestration engines.
  • βš™οΈ Core Infra: Architecting low-latency asynchronous task queues and high-throughput data pipelines.

🌟 Flagship Projects

⚑ PixelQueue

Vision Intelligence Infrastructure: A high-performance, async control panel for human-in-the-loop AI annotation.

Decoupled ML microservices platform eliminating UI bottlenecks with instant rendering and async model workers.

  • πŸš€ Asynchronous ML: Non-blocking auto-labeling via PyTorch, LayerCAM & YOLO.
  • ⚑ Zero-Latency UI: Hardware-accelerated React-Konva staging canvas.
  • πŸ”„ Decoupled Scale: Distributed worker orchestration via Celery message brokers.

"Finding the needle in the haystack, from 400ft above."

A coarse-to-fine computer vision pipeline for efficient small-object detection in high-resolution (2K/4K) aerial imagery.

  • 🎯 Smart Attention: Uses LayerCAM to localize semantic hotspots before inference.
  • ⚑ 80%+ Background Skipped: Intelligently zooms into regions of interest, bypassing empty tiles.
  • πŸ† High Efficiency: Outperforms blind sliding-window methods (SAHI) in speed and accuracy.

🎨 Neural Canvas

Transform any image into a masterpiece β€” in real-time.

Fast feed-forward neural style transfer generating stylized imagery in a single forward pass.

  • πŸš€ Real-Time Inference: Custom residual architecture trained with perceptual loss (VGG-16).
  • πŸ” Artifact-Free: Instance Normalization for high-fidelity texture synthesis.
  • πŸ“¦ Edge Ready: Full ONNX runtime export for edge deployment.

πŸ“¦ More Projects

🧭 pygog (Google CLI Agent)
CLI for Google Workspace (Gmail, Drive, Calendar) with built-in natural language AI agent support.
<Python> <Google APIs> <LLM Agents>

πŸ“ Depth Estimation + Semantic Seg.
Multi-modal depth completion (RGB + sparse depth + semantics) with NYU Depth v2 supervision.
<PyTorch> <NYU-Depth-v2> <Encoder-Decoder>

🌐 Open Source Contributions

Active contributor across AI infrastructure, agent frameworks, and computer vision libraries:

  • huggingface/optimum: Resolved CLI subcommand resolution for symlinked environments on RHEL/lib64 systems.
  • pydantic/pydantic-ai: Upgraded Anthropic code execution tooling integration.
  • lancedb/lancedb: Migrated Python Gemini embedding provider to google-genai SDK and resolved async event loop blocking in AsyncTable.add.
  • albumentations-team/AlbumentationsX: Implemented volumetric 3D transform noise models for high-performance CV augmentation.
  • SynapseKit/SynapseKit: Core Collaborator (100+ PRs) β€” Built native observability, VoiceAgent pipelines, persistent memory, and multimodal RAG ingestion.

πŸ“Š Telemetry

streak-stats

πŸ”— Connect & Explore

Website Β β€’Β  Searchlight Live App Β β€’Β  Email Me


Built by DhruvGarg111

Pinned Loading

  1. Neural-Style-Transfer Neural-Style-Transfer Public

    Python

  2. The-Searchlight-Protocol The-Searchlight-Protocol Public

    Python