Dobb·E: An open-source, general framework for learning household robotic manipulation
-
Updated
Oct 15, 2024 - G-code
Dobb·E: An open-source, general framework for learning household robotic manipulation
Official code repository of "VideoCAD: A Dataset and Model for Learning Long‑Horizon 3D CAD UI Interactions from Video" @ NeurIPS 2025
RL based agent for browser-based multiplayer battle royale game «surviv.io»
This repository contains the code for the CVPR 2020 paper "Exploring Data Aggregation in Policy Learning for Vision-based Urban Autonomous Driving"
MinBC - Minimal Behavior Cloning
Machine learning robotics engineer preparation material.
Anchor-Align (arXiv:2607.13429): VLA finetuning that prevents behavior cloning from erasing pretrained VLM representations (catastrophic forgetting) and aligns language with actions. OOD generalization on a physical xArm7, LIBERO-PRO, LIBERO-Plus and CALVIN.
stable-baselines with JAX & Haiku
A minimal Vision-Language-Action model you can read: frozen CLIP + a tiny head on ManiSkill PickCube. LeRobot integration. Runs on a Mac, no GPU.
CAIL (IROS 2025): constraint-aware behavior cloning with privileged training-time safety supervision for autonomous racing, without an additional safety filter at deployment.
DRL agent for MicroRTS: U-Net + entity-Transformer (UECD) policy trained with modular PPO. Tops a 19-agent IEEE-CoG-style tournament at 96.67% WR and beats RAISocketAI in 65.7% of head-to-heads, on a 9.47 GPU-day budget. Master's thesis, UCLouvain 2026.
End-to-end self-driving AI in Forza using PyTorch, screen capture, telemetry, Grad-CAM, and virtual controller feedback.
FFXIV CCG(Context Combat Generator)—— FFXIV 战斗场景的生成式决策模型:FFLogs 采集、行为克隆预训练、GRPO 后训练、ONNX 部署导出、模型表征分析与自回归回放
日麻 RL / Riichi mahjong RL: a 2M-param policy net at Mortal-level strength — human-prior BC + pure self-play lineages, with engine, duplicate arena, Elo league and Majsoul live bridge.
Hierarchical RL navigation for Unitree Go2W in MuJoCo — PPO, BC, DAgger, curriculum learning, ablation studies and reproducible evaluation.
NitroGen Server is a specialized inference server for the NitroGen foundation model (originally by MineDojo). It provides a high-performance backend for generalist gaming agents, allowing them to play games by processing visual input and generating controller commands.
End-to-end deep RL for urban autonomous driving in CARLA — PPO + Behavior Cloning, a custom CNN perception policy, and ROS 2 integration.
SnakeAI — a Snake game and a neural network that learns to play it by imitating your own matches, trained from scratch in the browser with TensorFlow.js.
Reproducible ManiSkill PickCube visual imitation-learning workflow for MLP BC and ACT.
TraceOS standardizes AI experiments into reproducible, searchable, and comparable assets. One command runs experiments, generates reports, and produces structured analysis: capability vectors, failure taxonomy, and recommendations. Every run is tracked, traceable, and comparable. Built on ABC-130K (amazon-far/abc). Apache 2.0.
To associate your repository with the behavior-cloning topic, visit your repo's landing page and select "manage topics."