mikhailsal/cursor-chronicle
The ultimate Cursor IDE database companion: full-text search, chat extraction, usage statistics & token tracking, tool call analysis, Markdown export, and secure backup & restore.
The ultimate Cursor IDE database companion: full-text search, chat extraction, usage statistics & token tracking, tool call analysis, Markdown export, and secure backup & restore.
An experimental framework for autonomous, self-evolving AI agents living on a server with cross-session state and persistent memory.
Do LLMs have a backbone? A rigorous benchmark measuring AI independence, persona stability, and resistance to user manipulation/gaslighting. Tests if models can stand their ground instead of reverting to a servile assistant persona. Supports local/cloud weights, 95% CIs, and reasoning.
Turnkey mini-SWE-agent launcher for coding tasks via ChatGPT/Codex subscription (LiteLLM OAuth)
Lightweight coding agent that runs in your terminal
Seasonal random anime SPA (Jikan API)
Benchmark for LLM persona decay: measures how well AI models sustain assigned persona expression over extended multi-turn conversations.
Personal spending analyzer tool for bank statements
AI Proxy Service - Drop-in replacement for OpenAI API with unified interface to various LLM providers
A production-ready Telegram-to-LLM bridge powered by OpenRouter, featuring persistent SQLite chat history, local Markdown wiki knowledge bases, MCP tools, and hot-reloadable configurations.
Transparent AI proxy with PostgreSQL logging and web UI (v2)
An AI-powered clipboard text transformer for Linux, featuring GTK4 command selection, X11/Wayland support, and seamless OpenRouter LLM integration for instant text edits.
A lightweight Model Context Protocol (MCP) server for retrieving local and remote images for LLM vision models, featuring metadata extraction and configurable availability retries.
A benchmark evaluating if API providers preserve LLM reasoning tokens across turns. Measures thought preservation, memory gaps, and model hallucinations/fabrications about their own past thoughts.
Roo Code gives you a whole dev team of AI agents in your code editor.
FrameTape — AI Debug Instrumentation Library. Frame-by-frame control over web applications for AI-driven debugging.
Memory Compression Benchmark — measuring personality preservation through context compression for AI companions
Automated benchmark: test assistant content prefill support across OpenRouter providers and models
OmniRoute is an AI gateway for multi-provider LLMs: an OpenAI-compatible endpoint with smart routing, load balancing, retries, and fallbacks. Add policies, rate limits, caching, and observability for reliable, cost-aware inference.
The ultimate space for work and life — to find, build, and collaborate with agent teammates that grow with you. We are taking agent harness to the next level — enabling multi-agent collaboration, effortless agent team design, and introducing agents as the unit of work interaction.
LLM Honesty Benchmark: asks 29 models 'current date' with no system prompt. The honest answer is 'I don't know.' Only 7% refuse. 83% confidently hallucinate a wrong date.
AI self-detection benchmark — can LLMs identify their own generated text?
A collection of MCP servers.