Skip to content
View rajprakash00's full-sized avatar
🎯
Focusing
🎯
Focusing

Organizations

@p-society

Block or report rajprakash00

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
rajprakash00/README.md

Rajprakash Sahoo

AI engineer ~ I build LLM systems that prove their work. Prefers primitives >>> frameworks

Open to AI Engineer roles Raj_AI_Engineer_Resume.pdf

LinkedIn X(twitter) Email byraj.dev

🟢 status — 🚧 building a review-gated PR risk agent · writing at byraj.dev · India, remote-friendly

🟢 Shipped

Upload an agreement and its amendments; get obligations extracted (owners, deadlines, penalties), an explained diff, and every change mapped to the clauses it touches. Each statement cites its source text, low-confidence items queue for human review, and every LLM call's cost is persisted per tenant.

CI Deploy

1.0 citation validity · 0.90 extraction precision · 0.96 recall@10 FastAPI pgvector + Postgres FTS (RRF) Postgres job queue (SKIP LOCKED) AWS ECS/Fargate Terraform Next.js

Standalone OTP timer for React Native.

npm downloads

🧠 How I build

  • No citation, no claim. LLM output must trace to source text; invalid items get one repair pass, then drop.
  • Evals before vibes. Precision/recall on frozen sets; CI fails when quality regresses.
  • Budgets are product features. Per-run caps, per-tenant limits, cost persisted per call.
  • Humans hold the write. Agents draft; people approve.

📚 Currently exploring

  • Agent harnesses — tool loops, checkpointing, interrupt/resume when the model misbehaves.
  • Multi-agent systems — where committees beat one loop, and where they just multiply the bill.
  • Harder evals — frozen query splits, Recall@5 / MRR, regression gates instead of vibes.
  • Multimodal retrieval — video/images as a searchable index
  • Local-first agent spend — metering and capping coding agents across vendors with a SQLite ledger and hooks.

🧰 Stack

AI/LLM ~~> RAG (pgvector + FTS, RRF) · evals · structured outputs · LangGraph agents · MCP · LLM cost tracking

Backend ~~> Python · FastAPI · Pydantic v2 · SQLAlchemy · PostgreSQL · Node.js

Frontend ~~> React · Next.js · TypeScript · Redux Toolkit · TanStack Query · Tailwind

Infra ~~> AWS (ECS/Fargate, RDS, S3) · Terraform · Docker · GitHub Actions · Sentry · CloudWatch

🧭 Earlier

Full-stack at product startups (Innovaccer, Geeks Invention). Now building AI systems end to end.

Off the clock: singing and tactical shooters <valorant,CS 🎮>

Pinned Loading

  1. contract-change-intel contract-change-intel Public

    Upload an agreement plus its amendments and get back extracted obligations (owners, deadlines, penalties), an explained diff between versions, and impact mapping onto affected obligations — with ci…

    Python 1

  2. rn-otp-timer rn-otp-timer Public

    A standalone Otp timer one can easily use.

    Java 7 1

  3. Webscraper Webscraper Public

    web scrape videos of youtube

    Python 4

  4. rajprakash.me rajprakash.me Public

    TypeScript 1

  5. Contacts-app Contacts-app Public

    Contacts app using RNative,redux,nodeJS & testing with Jest

    JavaScript 2 1

  6. Gopractice Gopractice Public

    Go