28 karma · joined January 24, 2018
Frisson Labs is building AI players that play like Discord friends. The hard part is making them respond with low latency, persistent personality, voice timing, and actual game competence: understanding what is happening, deciding what to do, and executing in real time.
Open roles:
- Founding ML Engineer - General: own the full stack for how AI companions play games: game-state understanding, learned policies/controllers, world models, action representation, MLP/policy heads, memory, evals, data pipelines, inference, and product iteration. Strong fit if you have shipped applied ML systems and want to own messy model behavior all the way from research prototype to live gameplay.
- Founding ML Engineer - Audio/Speech: explore and ship low-latency speech systems for real-time play: duplex/streaming voice models, speech understanding, turn-taking, interruption handling, latency/quality tradeoffs, prosody/emotion, and conversational evals. Strong fit if you have deep audio/speech ML experience and care about making voice feel socially present during gameplay.
Comp: - ML Engineer roles: $150K-$200K + 1.0-2.0% equity
Apply: email founders@frisson-labs.com with the role in the subject, plus your resume/GitHub/research/portfolio and a short note on why AI companions + gaming interests you.
Gemma can take in audio, images, and text, but only talks back in text. Mimi can turn codec tokens back into speech. So I froze both sides and trained a small graft in the middle: Gemma hidden states -> Mimi audio tokens.
I've enjoyed playing with this because the bad audio outputs have sounded hilarious