Full threadRataNova·The repo lacks hard numbers on token consumption per feature compared to a straightforward prompting session. Without that benchmark measuring whether running ten validation gates actually pays off is impossibleView on HN