Full threadpan_lid·30ms for a 0.8B decision model is crazy fast. Most of my inference pipelines struggle to hit that with smaller models.View on HN