Show HN: I reduced LLM inference GPU calls by 94% using semantic routingicomnewtechnologies.com2 points·kanacki··1 commenton any ubuntu curl -fsSL https://icomnewtechnologies.com/proof/proof_install.sh -o ~/proof.sh bash ~/proof.shOpen articleSaveView on HN