Okyumi/Okyumi.github.io
Yumi Omori's research and personal writing.
Doing research @nyuad. Lifelong sidequester :)
Yumi Omori's research and personal writing.
RL environment with exact I Wanna fangame physics: 5.8M steps/s C core, PufferLib-ready, goal-conditioned (HER/GCRL) support, 10 levels
StableCRL BuilderBench with permutation-aware DCC continual reinforcement learning.
Overland travel route planner — from Tokyo to Antarctica, no flights required.
Now, Stronger AI Pushes Frontiers, Stronger Our Shared Future.
Codebase of the paper "Demystifying the Mechanisms Behind Emergent Exploration in Goal-conditioned RL"
Official website for the China-Gulf Forum (CGF) — a student-led platform fostering China-Gulf dialogue since 2019 | china-gulf-forum.org
OpenClaw plugin — plan international overland routes by train, bus, and ferry. 191 cities · 506 connections · 103 operators · 22 countries · Dijkstra pathfinding
CLI-Anything: Making ALL Software Agent-Native
OverTrack — Overland international travel route planner. React Native (Expo) iOS app.
An Open-Source Asynchronous Coding Agent
Contrastive RL for Board Games: experiments applying InfoNCE-based GCRL to Chess, Go, and Connect-4
An API standard for single-agent reinforcement learning environments, with popular reference environments and related utilities (formerly Gym)
High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)
Fully autonomous & self-evolving research from idea to paper. Chat an Idea. Get a Paper. 🦞
Featured blog posts from omoriyumi.com
A fast, private, client-side PDF utility for splitting, merging, and compressing PDFs. No uploads, no servers — everything runs in your browser.
ENDWALKER — A parametric design sculpture exploring entropy, solar geometry, and generative form. Interactive project website by Yumi Omori, NYU Abu Dhabi.
Structured notes from PSYCH-UH 2412 Cognitive Neuroscience at NYU Abu Dhabi. Covers perception, attention, working memory, long-term memory, and consciousness.
Continual Knowledge Adaptation for Reinforcement Learning
🧑🏫 60+ Implementations/tutorials of deep learning papers with side-by-side notes 📝; including transformers (original, xl, switch, feedback, vit, ...), optimizers (adam, adabelief, sophia, ...), gans(cyclegan, stylegan2, ...), 🎮 reinforcement learning (ppo, dqn), capsnet, distillation, ... 🧠
High-quality single-file implementations of SOTA Offline and Offline-to-Online RL algorithms: AWAC, BC, CQL, DT, EDAC, IQL, SAC-N, TD3+BC, LB-SAC, SPOT, Cal-QL, ReBRAC