Skip to content
View rajj28's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report rajj28

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
rajj28/README.md
Terminal. whoami: Ruturaj Sonkamble, AI engineer, Pune. I build agents that act on real systems, then measure if they got it right. tail receipts.log: argus, 0 LLM calls over 10 runs and 5 of 5 hidden bugs caught; racelab, 0 of 50 bad commits versus 45 to 48 of 50 for blind retry; verdict, judge-bias correction raised rank agreement from tau 0.669 to 0.877; anomaly, 38 times faster than real time on a 6 GB laptop GPU.

SDE-1 at Concentrix Catalyst, building agentic workflows in production. On my own time I build systems where an agent's mistake would be expensive, and I publish how I checked them.

Receipts

Project What it is, and the proof
Argus
showcase
UI-testing agent: an LLM writes each test once, then it replays deterministically and heals UI changes
0 LLM calls over 10 runs · 5/5 hidden bugs caught
RaceLab
demo
What an agent should do when the data it reasoned about changes before it commits
0/50 bad commits vs 45–48/50 for blind retry, over 5,000 decisions
Continuity Release checks for dubbed and subtitled films; Grafana decides whether each market can ship
5 markets · 87 checks · agents fix assets, never the verdict
VERDICT
live
Self-hostable hackathon judging with judge-bias correction
Rank agreement τ 0.669 → 0.877 · 1,144 tests
Video anomaly
page
Traffic and CCTV anomaly detection across 11 classes, with event timestamps
38× faster than real time on a 6 GB laptop GPU
gpu-fleet-operator Go control plane for a GPU inference fleet (Kubernetes, Temporal)
56 tests that need no cluster to run

How I build

  • Measure, then claim. Every number above comes from a run you can repeat, like Argus's recorded results.
  • Keep the failures in the write-up. RaceLab lists the predictions it falsified, and the anomaly detector documents the model I didn't ship.
  • Agents read the verdict; they don't write it. In Continuity, agents reach Grafana through a write-disabled MCP server.

Also

1st place, DigiPay Pro (NPCI) competition, IIT Bombay Techfest 2024, out of 200+ teams (code)
Merged fixes in Eclipse Theia, Eclipse Thing-Web and rust-ffmpeg-sys
Writing: Teaching AI to read financial tables: fine-tuning LayoutLMv3

Python · Go · TypeScript · PyTorch · Playwright · FastAPI · Django · PostgreSQL · CockroachDB · Docker · GCP · AWS

ruturaj.is-a.dev · LinkedIn · ruturajsonkamble29@gmail.com

Pinned Loading

  1. FraudDetectionUsingGANs FraudDetectionUsingGANs Public

    Jupyter Notebook

  2. aerial-site-intelligence aerial-site-intelligence Public

    Python

  3. Autonomous-Drone-Fleet-Coordinator-Multi-Agent-System Autonomous-Drone-Fleet-Coordinator-Multi-Agent-System Public

    Architected a multi-agent framework using LangGraph and CrewAI where specialized agents (perception, planning, execution) autonomously coordinate drone fleet missions via agent-to-agent communicati…

    Python

  4. Fine-tuning-LayoutLMv3-for-Financial-Document-Table-Structure-Recognition Fine-tuning-LayoutLMv3-for-Financial-Document-Table-Structure-Recognition Public

    Python

  5. query-retrieval-using-RAG-pinecone-gpt-40-min query-retrieval-using-RAG-pinecone-gpt-40-min Public

    A production-ready Retrieval-Augmented Generation (RAG) system that intelligently adapts to multiple document domains while maintaining excellence in Insurance, Legal, HR, and Compliance as primary…

    Python

  6. -Multi-Drone-Routing-Optimization-Challenge -Multi-Drone-Routing-Optimization-Challenge Public

    Multi-Drone Routing Optimization Challenge - HackerRank × Goldman Sachs Hackathon (Score: 42.02/100)

    Python