Taylor Kolasinski
Founder of Poisson Labs. ML systems & research, reinforcement learning, edge AI. Previously Matter AI, Better.com, WeWork. Brooklyn.
Writing about reinforcement learning, large-scale model training, and simulation.
Notes
View all- Sep 24, 2026 145 of 783 — I measured whether Jev, TypeSafe's decision model, gives the same yes or no twice on the same request near its thresholds. Roughly one near-threshold decision in five flipped at 0.50 and 0.60 while the headline rate barely moved.
- Sep 15, 2026 45.0% and 44.0% — I ran the same agent benchmark twice with the same seeds. The two scores were 45.0% and 44.0%. Compared task by task, 44 of the 200 matched pairs came out differently, and all 200 took a different path.
- Sep 9, 2026 Where a 14-DoF Biped Falls — Frozen checkpoints of a 14-DoF walking policy, measured under lateral pushes with Proofload. The first answer was wrong, the map showed a cliff and a slope, and a retrain meant to widen the boundary made it worse in 14 of 14 cells.
Viz
View allVisual evidence from reinforcement learning, simulation, and large-scale model training.
Logs
View all- Sep 10, 2026 Proofload: A Frozen Test That Owned Too Much, and a Hole in the Ledger
- Sep 9, 2026 Bipedal Locomotion: Developmental Trajectory Audit and Wrap
- Sep 9, 2026 Bipedal Locomotion: The VelStand Intervention Trial
- Sep 9, 2026 Bipedal Locomotion: High-Frequency Telemetry and Mechanism
- Sep 9, 2026 Bipedal Locomotion: Resolving the Failure Boundary
- Sep 8, 2026 Bipedal Locomotion: The Gait-Phase Confound and Replication Audit
- Sep 8, 2026 Bipedal Locomotion: Specimen Sizing and the Initial Sweep
- Sep 6, 2026 Proofload: A Refusal That Offered an Override That Also Refused
- Sep 5, 2026 Proofload: The Stopping Rule Fired and the Third Arm Never Launched
- Sep 5, 2026 Proofload: The Pilot Could Not Be Launched
- Sep 4, 2026 Proofload: The Trace Spec Said It Was Not a Retrofit
- Sep 3, 2026 Proofload: Nine Modules and One Competing-Risks Bug
- Sep 2, 2026 Proofload: The Diff Reproduced a Published Result on the First Run
- Aug 22, 2026 Reward Latency: The Instrument Broke More Than the Subject
- Aug 18, 2026 Failure Boundary: Cold Starts, Host Pointers, and Hitting the Spend Limit
- Aug 18, 2026 Failure Boundary: The Retrain Moved It
- Aug 15, 2026 Failure Boundary: The Survivor That Can't Be Re-Simulated
- Aug 13, 2026 Failure Boundary: The Gate That Kept Failing
- Aug 13, 2026 Failure Boundary: 6,400 Rollouts in 126 Seconds
- Apr 2, 2026 Argus: Crack Detection Pipeline
- Apr 1, 2026 Argus: Bridge Inspection RL Training
- Mar 28, 2026 Stratum: Phase 2 Baseline Implementation
- Mar 28, 2026 Stratum: Cognitive Synthesis & Scientific Lockdown
- Mar 27, 2026 Stratum: Phase 2 Research Pivot
- Mar 25, 2026 Stratum: Bug Fixes & Game Mechanics
- Mar 24, 2026 Stratum: Design & Engine Build
- Mar 23, 2026 DeepSky ATC: MLP Scaling Limits Identified
- Mar 22, 2026 DeepSky ATC: MLP Baseline Refactor
- Mar 21, 2026 DeepSky ATC: 50-Agent Coordination Verified
- Mar 20, 2026 DeepSky ATC: Stage 4 Production Training
- Mar 15, 2026 DeepSky ATC: Reward Function Breakthrough
- Feb 12, 2026 Cable Mind: Vision Training & Video Pipeline
- Feb 11, 2026 Cable Mind: Camera Placement & Scene Design
- Feb 10, 2026 Cable Mind: UR5e Competition Environment
- Jan 21, 2026 MARL Cooperative Navigation: Geometry Problem & The Abyss
- Jan 18, 2026 MARL Cooperative Navigation: Reward Hacking & Phase 1 Solved
- Jan 17, 2026 MARL Cooperative Navigation: Reward Shaping and Local Optima
- Jan 16, 2026 MARL Cooperative Navigation: Gridworld to MAPPO
- Jan 15, 2026 DeepSeek mHC: The Bomb
- Jan 14, 2026 DeepSeek mHC: Infrastructure Hell
- Jan 10, 2026 DeepSeek mHC: The Stream Persistence Bug
- Jan 9, 2026 Cable Insertion RL: Coordinate Frame Bugs & Reward Ablations
- Jan 9, 2026 mHC: Depth Sweep Validation
- Jan 8, 2026 Cable Insertion RL: MuJoCo Environment Design
- Jan 7, 2026 mHC: Stream Persistence Fix
- Jan 6, 2026 mHC: Sinkhorn-Knopp Implementation
- Jan 1, 2026 VL-JEPA Reproduction
- Dec 15, 2025 Edge AI Platform Architecture