Dobb·E: An open-source, general framework for learning household robotic manipulation
-
Updated
Oct 15, 2024 - G-code
Dobb·E: An open-source, general framework for learning household robotic manipulation
Official code repository of "VideoCAD: A Dataset and Model for Learning Long‑Horizon 3D CAD UI Interactions from Video" @ NeurIPS 2025
RL based agent for browser-based multiplayer battle royale game «surviv.io»
This repository contains the code for the CVPR 2020 paper "Exploring Data Aggregation in Policy Learning for Vision-based Urban Autonomous Driving"
MinBC - Minimal Behavior Cloning
Machine learning robotics engineer preparation material.
Anchor-Align (arXiv:2607.13429): VLA finetuning that prevents behavior cloning from erasing pretrained VLM representations (catastrophic forgetting) and aligns language with actions. OOD generalization on a physical xArm7, LIBERO-PRO, LIBERO-Plus and CALVIN.
stable-baselines with JAX & Haiku
A minimal Vision-Language-Action model you can read: frozen CLIP + a tiny head on ManiSkill PickCube. LeRobot integration. Runs on a Mac, no GPU.
CAIL (IROS 2025): constraint-aware behavior cloning with privileged training-time safety supervision for autonomous racing, without an additional safety filter at deployment.
DRL agent for MicroRTS: U-Net + entity-Transformer (UECD) policy trained with modular PPO. Tops a 19-agent IEEE-CoG-style tournament at 96.67% WR and beats RAISocketAI in 65.7% of head-to-heads, on a 9.47 GPU-day budget. Master's thesis, UCLouvain 2026.
End-to-end self-driving AI in Forza using PyTorch, screen capture, telemetry, Grad-CAM, and virtual controller feedback.
日麻 RL / Riichi mahjong RL: a 2M-param policy net at Mortal-level strength — human-prior BC + pure self-play lineages, with engine, duplicate arena, Elo league and Majsoul live bridge.
Hierarchical RL navigation for Unitree Go2W in MuJoCo — PPO, BC, DAgger, curriculum learning, ablation studies and reproducible evaluation.
NitroGen Server is a specialized inference server for the NitroGen foundation model (originally by MineDojo). It provides a high-performance backend for generalist gaming agents, allowing them to play games by processing visual input and generating controller commands.
End-to-end deep RL for urban autonomous driving in CARLA — PPO + Behavior Cloning, a custom CNN perception policy, and ROS 2 integration.
SnakeAI — a Snake game and a neural network that learns to play it by imitating your own matches, trained from scratch in the browser with TensorFlow.js.
TraceOS standardizes AI experiments into reproducible, searchable, and comparable assets. One command runs experiments, generates reports, and produces structured analysis: capability vectors, failure taxonomy, and recommendations. Every run is tracked, traceable, and comparable. Built on ABC-130K (amazon-far/abc). Apache 2.0.
Reproducible ManiSkill PickCube visual imitation-learning workflow for MLP BC and ACT.
宝可梦卡牌 PTCG 对战 AI 竞技方案(Silver-medal)—— 纯 JAX 规则引擎 + BC(行为克隆) + PPO 强化学习,含 8×A100 分布式训练与 PPO 训练池构成。
To associate your repository with the behavior-cloning topic, visit your repo's landing page and select "manage topics."