Posts by Tags

3D Gaussian Splatting

3D Geometry

3D Hand Reconstruction

3D Point Cloud

3D Reconstruction

3D Scene Flow

3D Scene Generation

3D Vision

6D Pose Estimation

A100

AI

Bootstrapping and the Game of Entrepreneurship

11 minute read

Published:

Bootstrapping is the same pattern across compilers, AI, and startups: borrow external structure to cold-start, then recursively replace dependencies until the system sustains itself.

The Singularity is Near

1 minute read

Published:

Autonomous agents are now cheap and networked enough to spread faster than any institution can contain — the question is what order emerges after control becomes partial.

Intelligence Is Not Only Reasoning

9 minute read

Published:

Intelligence depends on retrieval and communication as much as reasoning — a powerful model limited by information gaps is still limited.

Cognitive Bandwidth in the AI Agent Era

8 minute read

Published:

As AI tools get more capable, the bottleneck shifts from tool friction to human cognitive bandwidth — compressing intent and steering effectively becomes the key skill.

AI Agents

The Singularity is Near

1 minute read

Published:

Autonomous agents are now cheap and networked enough to spread faster than any institution can contain — the question is what order emerges after control becomes partial.

Intelligence Is Not Only Reasoning

9 minute read

Published:

Intelligence depends on retrieval and communication as much as reasoning — a powerful model limited by information gaps is still limited.

Cognitive Bandwidth in the AI Agent Era

8 minute read

Published:

As AI tools get more capable, the bottleneck shifts from tool friction to human cognitive bandwidth — compressing intent and steering effectively becomes the key skill.

AI Ethics

AI for Science

Action Chunking

Action Coherence

Action Representation

Actuator Modeling

Admittance Control

Force Control and the Missing Layer in Embodied Intelligence

11 minute read

Published:

A late-night conversation over grilled skewers revealed a gap: most embodied AI researchers don’t think about what the robot arm is actually doing at the control level — and closing that gap might reshape how we think about action spaces.

Affordance Learning

Agentic AI

Agentic Robotics

Alignment

Apple Silicon

Articulated Objects

Articulated Tools

Asynchronous Inference

Automated Discovery

Autonomous Research

Behavior Cloning

Behavior Priors

Benchmark

Benchmarking

Bimanual Interaction

Bimanual Manipulation

Blind Grasping

Blog Notes

Book Notes

Bootstrapping

Bootstrapping and the Game of Entrepreneurship

11 minute read

Published:

Bootstrapping is the same pattern across compilers, AI, and startups: borrow external structure to cold-start, then recursively replace dependencies until the system sustains itself.

CEM

CMA-ES

CUDA

Chain of Thought

Claude Code

Claude Skills

CoRL

[Paper Notes] DayDreamer: World Models for Physical Robot Learning - CoRL 2022

1 minute read

Published:

Key information

  • This paper learns two models: a world model trained on off-policy sequences through supervised learning, and an actor-critic model to learn behaviors from trajectories predicted by the learned model.
  • The data collection and learning updates are decoupled, enabling fast training without waiting for the environment. A learner thread continuously trains the world model and actor-critic behavior, while an actor thread in parallel computes actions for environment interaction.

Code Generation

Codex

I Vibe Coded an Operating System

17 minute read

Published:

I built a scripting-first OS prototype in a weekend with Codex, where the system language gradually takes over its own environment from the host.

Coding Agents

Cognition

Cognitive Bandwidth in the AI Agent Era

8 minute read

Published:

As AI tools get more capable, the bottleneck shifts from tool friction to human cognitive bandwidth — compressing intent and steering effectively becomes the key skill.

Cognitive Science

Communication

Intelligence Is Not Only Reasoning

9 minute read

Published:

Intelligence depends on retrieval and communication as much as reasoning — a powerful model limited by information gaps is still limited.

Compliance Control

Computer Vision

Computing

Consciousness

Contact Modeling

Contact-Rich Manipulation

Continual Learning

Continuous Control

Contrastive Learning

Coordinate Frames

Cross-Embodiment

Cross-Embodiment Learning

Cross-Embodiment Transfer

DINO

Data Augmentation

Data Curation

Data-Centric AI

Dataset

Deformable Manipulation

Demonstration Retrieval

Depth Reconstruction

Design Optimization

Dexterous Grasping

Dexterous Hands

Dexterous Manipulation

Differentiable Simulation

Diffusion Language Models

Diffusion Models

Diffusion Policy

Digital Teleoperation

Digital Twins

DiscoRL

Discrete Diffusion

Distance Fields

Distillation

Distributed Training

Dynamic Manipulation

Dynamics Models

EMG

Economics

The Singularity is Near

1 minute read

Published:

Autonomous agents are now cheap and networked enough to spread faster than any institution can contain — the question is what order emerges after control becomes partial.

Egocentric Data

Egocentric Learning

Egocentric Video

Egocentric Vision

Embodied AI

Embodied Agents

Embodied Intelligence

Force Control and the Missing Layer in Embodied Intelligence

11 minute read

Published:

A late-night conversation over grilled skewers revealed a gap: most embodied AI researchers don’t think about what the robot arm is actually doing at the control level — and closing that gap might reshape how we think about action spaces.

Embodied Planning

End Effectors

Energy-Efficient Locomotion

Engineering Notes

A Practical Map of LLM Training Ecosystems

8 minute read

Published:

A compact field note on how Megatron, Hugging Face, PyTorch, vLLM, SGLang, and verl fit together in a practical LLM training pipeline.

Entrepreneurship

Bootstrapping and the Game of Entrepreneurship

11 minute read

Published:

Bootstrapping is the same pattern across compilers, AI, and startups: borrow external structure to cold-start, then recursively replace dependencies until the system sustains itself.

Epistemology

Equivariance

Event Camera

Evolutionary Robotics

Explainable AI

FSDP

Fault-Tolerant Control

Flow Matching

Force Control

Force Control and the Missing Layer in Embodied Intelligence

11 minute read

Published:

A late-night conversation over grilled skewers revealed a gap: most embodied AI researchers don’t think about what the robot arm is actually doing at the control level — and closing that gap might reshape how we think about action spaces.

Force-Aware Manipulation

Force-aware Retargeting

Foundation Model

Foundation Models

FoundationPose

Functional Grasping

Games

Gaussian Splatting

Generalist Policy

Generative Models

Goal-Conditioned RL

Goal-conditioned

Graph Grammar

Graph Neural Networks

Grasp Synthesis

Grasping

Habitat

Hand Morphology

Hand Motion

Hand Pose Estimation

Hand Retargeting

Hand-Object Interaction

Hugging Face

A Practical Map of LLM Training Ecosystems

8 minute read

Published:

A compact field note on how Megatron, Hugging Face, PyTorch, vLLM, SGLang, and verl fit together in a practical LLM training pipeline.

Human Data

Human Demonstration

Human Demonstrations

Human Priors

Human Video

Human Video Learning

Human Videos

Human-Object Interaction

Human-Robot Interaction

Human-Robot Transfer

Human-in-the-Loop

Human-to-Robot Transfer

Humanoid

Humanoid Control

Humanoid Robotics

Humanoid Robots

Hybrid Systems

Imitation Learning

Impedance Control

Force Control and the Missing Layer in Embodied Intelligence

11 minute read

Published:

A late-night conversation over grilled skewers revealed a gap: most embodied AI researchers don’t think about what the robot arm is actually doing at the control level — and closing that gap might reshape how we think about action spaces.

In-Context Learning

In-Hand Assembly

In-Hand Manipulation

In-Hand Reorientation

In-Hand Rotation

Inference Acceleration

Information Retrieval

Intelligence Is Not Only Reasoning

9 minute read

Published:

Intelligence depends on retrieval and communication as much as reasoning — a powerful model limited by information gaps is still limited.

Information Theory

Institutions

Instruction Following

Intelligence

Intelligence Is Not Only Reasoning

9 minute read

Published:

Intelligence depends on retrieval and communication as much as reasoning — a powerful model limited by information gaps is still limited.

Internet

The Singularity is Near

1 minute read

Published:

Autonomous agents are now cheap and networked enough to spread faster than any institution can contain — the question is what order emerges after control becomes partial.

Inverse Dynamics

Inverse Kinematics

Isaac Sim

KV Cache

Kinematics

Knowledge

Knowledge Transfer

LLM

LLM Agents

LLM Inference

LLM Training

A Practical Map of LLM Training Ecosystems

8 minute read

Published:

A compact field note on how Megatron, Hugging Face, PyTorch, vLLM, SGLang, and verl fit together in a practical LLM training pipeline.

LLM Workflows

Language-Action Pre-Training

Large Language Models

Latent Action

Latent Actions

Latent Dynamics

LeRobot

Legged Locomotion

Legged Robots

Loco-Manipulation

Long-Context Policy

Long-Horizon Control

Long-Horizon Manipulation

Machine Learning Theory

Manipulability

Manipulation

Mathematics

Megatron

A Practical Map of LLM Training Ecosystems

8 minute read

Published:

A compact field note on how Megatron, Hugging Face, PyTorch, vLLM, SGLang, and verl fit together in a practical LLM training pipeline.

Memory

Memory Systems

Memory-Augmented Policy

Meta-Learning

Model Predictive Control

Model-Based Control

Model-Based RL

Model-Based Reinforcement Learning

Model-Free RL

Motion Capture

Motion Control

Motion Generation

Motion Imitation

Motion Retargeting

Motion Tracking

MuJoCo

Multi-Embodiment Learning

Multi-Fingered Hands

Multi-Object Grasping

Multimodal Learning

Multitask Learning

Music

Neuroscience

Novel View Synthesis

Object Tracking

Offline RL

On-Policy Data

Online RL

Open Source

Open-Vocabulary Robotics

Operating Systems

Bootstrapping and the Game of Entrepreneurship

11 minute read

Published:

Bootstrapping is the same pattern across compilers, AI, and startups: borrow external structure to cold-start, then recursively replace dependencies until the system sustains itself.

I Vibe Coded an Operating System

17 minute read

Published:

I built a scripting-first OS prototype in a weekend with Codex, where the system language gradually takes over its own environment from the host.

Optimal Experiment Design

Optimal Experimental Design

Optimization

PAL Robotics

PaliGemma

Paper Notes

Persona

Personal Thoughts

Force Control and the Missing Layer in Embodied Intelligence

11 minute read

Published:

A late-night conversation over grilled skewers revealed a gap: most embodied AI researchers don’t think about what the robot arm is actually doing at the control level — and closing that gap might reshape how we think about action spaces.

The Singularity is Near

1 minute read

Published:

Autonomous agents are now cheap and networked enough to spread faster than any institution can contain — the question is what order emerges after control becomes partial.

Intelligence Is Not Only Reasoning

9 minute read

Published:

Intelligence depends on retrieval and communication as much as reasoning — a powerful model limited by information gaps is still limited.

Cognitive Bandwidth in the AI Agent Era

8 minute read

Published:

As AI tools get more capable, the bottleneck shifts from tool friction to human cognitive bandwidth — compressing intent and steering effectively becomes the key skill.

Philosophy

Photometric Stereo

Physics

Physics Simulation

Physics-Based Animation

Pinocchio

Planning

Point Clouds

Policy Distillation

Policy Improvement

Policy Learning

Post-Training

Power

Pretraining

Privileged Learning

Productivity

Cognitive Bandwidth in the AI Agent Era

8 minute read

Published:

As AI tools get more capable, the bottleneck shifts from tool friction to human cognitive bandwidth — compressing intent and steering effectively becomes the key skill.

Programming Languages

I Vibe Coded an Operating System

17 minute read

Published:

I built a scripting-first OS prototype in a weekend with Codex, where the system language gradually takes over its own environment from the host.

Progress

Project Notes

I Vibe Coded an Operating System

17 minute read

Published:

I built a scripting-first OS prototype in a weekend with Codex, where the system language gradually takes over its own environment from the host.

Project SuperDex

Prompt Engineering

PyTorch

A Practical Map of LLM Training Ecosystems

8 minute read

Published:

A compact field note on how Megatron, Hugging Face, PyTorch, vLLM, SGLang, and verl fit together in a practical LLM training pipeline.

RGB Perception

RL Algorithms

RSS

Real-Time Control

Real-World RL

Rectified Flow

Reference Tracking

Reinforcement Learning

A Practical Map of LLM Training Ecosystems

8 minute read

Published:

A compact field note on how Megatron, Hugging Face, PyTorch, vLLM, SGLang, and verl fit together in a practical LLM training pipeline.

[Paper Notes] DayDreamer: World Models for Physical Robot Learning - CoRL 2022

1 minute read

Published:

Key information

  • This paper learns two models: a world model trained on off-policy sequences through supervised learning, and an actor-critic model to learn behaviors from trajectories predicted by the learned model.
  • The data collection and learning updates are decoupled, enabling fast training without waiting for the environment. A learner thread continuously trains the world model and actor-critic behavior, while an actor thread in parallel computes actions for environment interaction.

Representation Learning

Research

Research Automation

Research Notes

Residual Policy Learning

Retargeting

Reward Model

Reward Modeling

Reward Models

Rigid Dynamics

Robot Assembly

Robot Calibration

Robot Co-Design

Robot Control

Robot Design

Robot Dynamics

Robot Foundation Models

Robot Imitation Learning

Robot Learning

Robot Manipulation

Robot Navigation

Robot Pre-Training

Robot-Free Data Collection

Robotic Hands

Robotic Manipulation

Robotics

Force Control and the Missing Layer in Embodied Intelligence

11 minute read

Published:

A late-night conversation over grilled skewers revealed a gap: most embodied AI researchers don’t think about what the robot arm is actually doing at the control level — and closing that gap might reshape how we think about action spaces.

[Paper Notes] DayDreamer: World Models for Physical Robot Learning - CoRL 2022

1 minute read

Published:

Key information

  • This paper learns two models: a world model trained on off-policy sequences through supervised learning, and an actor-critic model to learn behaviors from trajectories predicted by the learned model.
  • The data collection and learning updates are decoupled, enabling fast training without waiting for the environment. A learner thread continuously trains the world model and actor-critic behavior, while an actor thread in parallel computes actions for environment interaction.

Robotics Data

Robotics Hardware

Robustness

Rolling Contact

Runtime Systems

I Vibe Coded an Operating System

17 minute read

Published:

I built a scripting-first OS prototype in a weekend with Codex, where the system language gradually takes over its own environment from the host.

Rust

SE(3)

SGLang

A Practical Map of LLM Training Ecosystems

8 minute read

Published:

A compact field note on how Megatron, Hugging Face, PyTorch, vLLM, SGLang, and verl fit together in a practical LLM training pipeline.

Sample Efficiency

Scaling Laws

Science

Science Robotics

Scientific Discovery

Self-Contact

Self-Hosting

Bootstrapping and the Game of Entrepreneurship

11 minute read

Published:

Bootstrapping is the same pattern across compilers, AI, and startups: borrow external structure to cold-start, then recursively replace dependencies until the system sustains itself.

Sensor Design

ShadowHand

Shared Autonomy

Sim-to-Real

Sim2Real

Simulation

Simulation Pre-training

Single-Image Reconstruction

Skill Chaining

Skill Composition

Soft Bodies

Soft Robotics

Spatial Grounding

Spatial Reasoning

State Representation

Strategy

Surface Electromyography

Surveillance

Symbolic Regression

Synthetic Data

System Identification

Systems

Tabletop Scenes

Tacit Knowledge

Tactile Feedback

Tactile Reasoning

Tactile Sensing

Tactile Simulation

Task and Motion Planning

Task-Oriented Manipulation

Technical Notes

Teleoperation

Test-Time Execution

Test-Time Training

Test-time Guidance

Theory

Tokenization

Tokenizer

Training Stages

Trajectory Optimization

Transformers

URDF

URMA

Unified Action Space

Unsupervised RL

VLA

VLM Agents

VLM Fine-tuning

Value Functions

Value Models

Video Diffusion

Video Foundation Models

Video Generation

Video World Models

Vision Language Action

Vision-Based Tactile Sensors

Vision-Language Models

Vision-Language-Action

Vision-Language-Action Models

Vision-Tactile Learning

Vision-and-Language Navigation

Visual Planning

Visual Representation Learning

Visual-Tactile Learning

Visuo-Tactile

Visuo-Tactile Learning

Visuotactile Learning

Visuotactile Sensing

Visuotactile Simulation

Wearable Sensing

Whole-Body Control

Whole-Body Manipulation

World Action Models

World Model

World Models

Zero-shot Generalization

Zero-shot Manipulation

update

Updating website

less than 1 minute read

Published:

A major site update done mostly by Codex — page refactoring, content restructuring, and bilingual support.

vLLM

A Practical Map of LLM Training Ecosystems

8 minute read

Published:

A compact field note on how Megatron, Hugging Face, PyTorch, vLLM, SGLang, and verl fit together in a practical LLM training pipeline.

website

Updating website

less than 1 minute read

Published:

A major site update done mostly by Codex — page refactoring, content restructuring, and bilingual support.