Available for new opportunities

Shaktisinh Chavda

AI Engineer • LLM Fine-Tuning • AI Agents & Evaluation

Building and evaluating intelligent systems that learn, reason, and act.

About

Building Intelligent Systems with Purpose

AI engineer and final-year B.Tech student with hands-on experience in LLM fine-tuning and reinforcement learning, AI agents, and model evaluation. Currently creating and reviewing AI benchmark tasks at Ambiguity Labs, while building practical AI tools for browser automation, web editing, and data analysis.

LLM Fine-Tuning & Reinforcement Learning
AI Evaluation & Benchmarking
AI Agents & Multi-Agent Systems
B.Tech AI/ML — LDCE, CGPA: 7.9/10
100%
BrowserGym Test Accuracy
2
AI Internships
4 GB
VRAM Training Setup
SC
AI Engineer
Anime avatar

Shaktisinh Chavda

B.Tech, AI & Machine Learning

LD College of Engineering · Ahmedabad

Technical Expertise

Core technologies and frameworks I work with

Languages

Python
SQL
JavaScript
TypeScript

ML / Deep Learning

PyTorch
TensorFlow
scikit-learn
HuggingFace Transformers
OpenCV

AI / LLM / GenAI

Large Language Models
LangChain
LangGraph
RAG
Fine-Tuning (SFT, LoRA, QLoRA)
GRPO
AI Evaluation & Benchmarking
Multi-Agent Systems
CrewAI
Google ADK

Tools / Infrastructure

FastAPI
Docker
Git
GitHub Actions
FAISS
ChromaDB
PostgreSQL
Ollama

Featured Projects

A selection of AI & ML projects I've built

Browser Control Agent with RL

Autonomous Web Navigation via Reinforcement Learning

Fine-tuned Qwen2.5-0.5B for web navigation using SFT followed by GRPO, raising accuracy from a 25% zero-shot baseline to 100% on the project's BrowserGym test suite. Built the Docker training setup and custom reward functions to run on a single 4 GB VRAM GPU.

PyTorchHuggingFace TRLLoRAGRPOBrowserGymDocker
View on GitHub

WebMorph AI

AI-Powered Visual Web Editor

Built a visual web editor that turns plain-English UI requests into live page updates using Gemini Flash or local models through Ollama. Created a CSS patching system that changes only the requested styles without breaking the page layout or interactive elements.

Next.jsFastAPIPlaywrightGemini APIOllamaTypeScript
View on GitHub

VizDataAI

Natural Language Data Analytics Platform

Built a multi-agent analytics tool that takes a plain-English question and handles data queries, transformations, and chart generation. Added local model support through Ollama so sensitive data can be analyzed offline.

FastAPILangChainMulti-Agent SystemsPandasMatplotlib
View on GitHub

Deepfake Detection Research

Cross-Domain Manipulated Media Detection

Deep learning framework for manipulated media detection using domain-adversarial training, improving cross-domain accuracy by 15% on unseen data distributions. Curated 10,000+ sample dataset spanning FaceSwap, Face2Face, and NeuralTextures generation techniques for robust evaluation.

TensorFlowGANsDomain-Adversarial NetworksOpenCV
View on GitHub

Experience

Professional journey in AI & Machine Learning

AI Research Intern

Ambiguity Labs

Remote

June 2026 – Present
  • Create and submit benchmark tasks used to test AI systems, with clear instructions and measurable pass/fail criteria.
  • Review tasks from other contributors to identify unclear requirements, incorrect results, inconsistent difficulty, and evaluation gaps.
AI EvaluationBenchmarkingTask DesignQuality Review

AI Engineer Intern

Unada Pvt. Ltd.

Ahmedabad

March 2026 – April 2026
  • Improved AI-generated construction progress reports through context engineering and replaced hardcoded report parameters with configurable settings.
  • Created image and video assets for client work using prompt engineering across multiple generative AI tools.
Context EngineeringPrompt EngineeringGenerative AIConfiguration Design

Education

Academic foundation in AI & Machine Learning

B.Tech, Artificial Intelligence & Machine Learning

L.D. College of Engineering, Ahmedabad

2023 – 2027CGPA: 7.9/10

Relevant Coursework

Deep LearningNatural Language Processing (NLP)Computer VisionReinforcement LearningData Structures & AlgorithmsDatabase Management Systems

Achievements

Key milestones and accomplishments

100% BrowserGym Benchmark

Improved Qwen2.5-0.5B from a 25% zero-shot baseline to 100% accuracy on the project's BrowserGym test suite.

AI Benchmark Contributor

Creates and reviews measurable benchmark tasks for evaluating AI systems at Ambiguity Labs.

Efficient RL Training

Built a Dockerized SFT and GRPO training setup that runs on a single 4 GB VRAM GPU.

Get in Touch

Have a project in mind or want to collaborate? Let's connect!

I'm always excited to discuss new opportunities, AI research collaborations, or interesting projects. Feel free to reach out through any of the channels below.