Vidit-Ostwal/Vidit-Ostwal
Keeps the read me updated
Always Getting Better
Keeps the read me updated
Framework for orchestrating role-playing, autonomous AI agents. By fostering collaborative intelligence, CrewAI empowers agents to work together seamlessly, tackling complex tasks.
TypeSafe Jev plays Flappy Bird via flappy-bird-gymnasium, spectated live over WebSocket/FastAPI
Personal Portfolio Website
Contains code for apex agents docker image
SGLang is a high-performance serving framework for large language models and multimodal models.
Muesli - local meeting transcription + dictation for macOS (Granola + WisprFlow alternative)
Ice-sliding maze environment built on OpenEnv
This application is a Text-to-SQL Chatbot designed to answer natural language questions about pharmaceutical sales and prescription data.
The RLM Interactive Console is a full-stack application designed to demonstrate and interact with Reinforcement Learning Models (or similar agentic systems).
MakeMyDocsBot is a smart documentation synchronization bot designed to help maintainers keep multi-language documentation up-to-date across feature branches.
An autonomous vision-language agent that plays Survival 3D on YouTube Playables. It watches the game through screenshots, sets goals from the UI, plans movement, and executes keyboard actions in a loop.
The harness drives a real browser (Playwright), explores the app via BFS, converts discovered paths into plain-English test goals, executes them in parallel, and verifies each step with an LLM-based verifier to produce structured findings and an HTML report.
An OpenEnv RL environment where an LLM agent plays the buyer and negotiates against an LLM-powered seller over real marketplace listings.
An interface library for RL post training with environments.
A skill that generates a rich, self-contained HTML report explaining any Prime Intellect verifiers environment.
Environments for LLM Reinforcement Learning
Modal-style sandbox API on top of Hugging Face Jobs
Training-Ready RL Environments + Evals
A verifiers-based multi-turn buyer-seller negotiation environment where a trainable buyer policy learns strategic bargaining against a fixed seller model using utility-, anchoring-, concession-, and formatting-based rewards.
This is majorly for my own learning purpose.
Train transformer language models with reinforcement learning.
Utils for Unsloth https://github.com/unslothai/unsloth
GitHub's official MCP Server
Codebase of all the personal blogs.
Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends
This is an attempt to get started with Golang
An open-source AI agent that brings the power of Gemini directly into your terminal.
An open protocol enabling communication and interoperability between opaque agentic applications.