Thanks to visit codestin.com
Credit goes to github.com

Skip to content
View sksanjoo2's full-sized avatar
๐Ÿ‡ฎ๐Ÿ‡ณ
Focusing
๐Ÿ‡ฎ๐Ÿ‡ณ
Focusing

Block or report sksanjoo2

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please donโ€™t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this userโ€™s behavior. Learn more about reporting abuse.

Report abuse
sksanjoo2/README.md

Hi, I'm Sanjeev ๐Ÿ‘‹

About Me

I build practical generative AI solutions and equally focus on evaluating whether those systems are reliable, safe, and useful in the real world.

Current Focus:

  • ๐Ÿงช LLM Evaluation & Responsible AI - Testing safety, hallucination, bias, guardrails
  • ๐ŸŽ™๏ธ Voice & Agentic AI - Building and evaluating voice agents
  • ๐Ÿ”„ Agentic Systems - Long-running, autonomous AI agents
  • ๐ŸŒ Multilingual/Indic AI - Voice, evaluation, synthetic data for Indian languages

๐Ÿš€ The Idea I'm Excited About

AI Evaluation & Testing Platform for Voice & Agentic AI

An automated platform that:

  • ๐ŸŽญ Generates realistic personas and multi-turn scenarios
  • ๐Ÿค– Interacts with AI agents to test them rigorously
  • ๐Ÿ›ก๏ธ Evaluates: safety, hallucination, instruction-following, bias, guardrails
  • ๐Ÿ“Š Produces evidence-based evaluation reports
  • ๐Ÿ† Uses LLM-as-a-Judge for intelligent evaluation

๐Ÿ’ก What I Can Contribute

  • GenAI/LLM architecture and optimization
  • Evaluation frameworks & metrics
  • Responsible AI & safety testing
  • Python & experimentation
  • Building realistic test scenarios

๐Ÿ”ง Tech Stack

AI/ML: Claude, GPT, LangChain, AWS Bedrock, RAG, Vector DBs
Voice: Speech Recognition, TTS, Voice Cloning
Data: Synthetic Data Generation, Multilingual NLP
Backend: Python, FastAPI, Jupyter

๐Ÿค Looking for Hackathon Teammates

I'm seeking builders and experimenters with complementary strengths:

  • ๐Ÿ—๏ธ Full-Stack/Product Development - Build UI, infrastructure, deployment
  • ๐Ÿค– AI/ML & Agents - Implement evaluation logic, agent orchestration
  • ๐ŸŽ™๏ธ Speech/Voice AI - Voice evaluation, TTS, speech processing
  • ๐ŸŽจ UI/UX Design - Make evaluation reports beautiful and actionable

What I value:

  • โœ… Shipping code over endless discussions
  • โœ… Iterating quickly and learning fast
  • โœ… Challenging assumptions
  • โœ… Curiosity and experimentation

๐Ÿ“ Featured Projects

Real-time voice AI agent evaluation framework. Tests guardrails, compliance, safety, and edge cases.

Advanced RAG system with multi-turn reasoning and evaluation capabilities.

Integration utilities for AWS Bedrock LLMs.

๐Ÿ”ด Hackathon Status

ACTIVELY SEEKING TEAMMATES FOR AI EVALUATION PLATFORM

If you're interested in building an AI evaluation platform and have complementary skills, let's connect!

๐Ÿ“ซ Let's Build Together

  • Open to collaboration on AI evaluation, voice agents, Indic AI
  • Excited about quick iterations and shipping working prototypes
  • Looking for curious builders, not idea discussers

"The best way to predict the future is to build it." โ€” Let's build something great together! ๐Ÿš€

Pinned Loading

  1. Nova-LLM-RAG-Agent Nova-LLM-RAG-Agent Public

    Jupyter Notebook

  2. Voice-Collection-Agent-Testing Voice-Collection-Agent-Testing Public

    Python

  3. llmasajudge llmasajudge Public

    Python