HNHacker News
TopNewBestAskShowJobs

_josh_meyer_

438 karma · joined January 3, 2022

building
submissionscomments
_josh_meyer_··on VAmoS Bench: Voice Agent Simulation Benchmark
hi there o/

i'm josh, the first author. been cooking this up for a while and happy to answer questions / comments!

_josh_meyer_··on Apple Foundation Models
the github repo: https://github.com/anthropics/ClaudeForFoundationModels
_josh_meyer_··on Amdahl's Law in Software Engineering
an optimistic rebuttal to Amdahl's law: https://lawsofsoftwareengineering.com/laws/gustafsons-law/
_josh_meyer_··on Amdahl's Law in Software Engineering
the original 1967 paper: https://dl.acm.org/doi/epdf/10.1145/1465482.1465560
_josh_meyer_··on BYU's Supermileage vehicle: Squeezing 2,145 miles out of a single gallon of fuel
local news video: https://www.youtube.com/watch?v=tz84xF6tn-M
_josh_meyer_··on The last six months in LLMs in five minutes
a 5-minute video version (with local TTS model) https://tldr-api.manatee.work/v/dmYg0U
_josh_meyer_··on LLM-as-a-Judge is asking the wrong question
icymi, here's the claude skill: https://github.com/veris-ai/veris-skills
_josh_meyer_··on LLM-as-a-Judge is asking the wrong question
Author here: Looking forward to discussion!
_josh_meyer_··on Nw_wrld is an event-driven sequencer for triggering visuals [video]
nw_wrld is an event-driven sequencer for triggering visuals using web technologies. It enables users to scale up audiovisual compositions for prototyping, demos, exhibitions, and live performances. Users code their own visual modules, then orchestrate them using the project's native UI composer.
_josh_meyer_··on GPT-5.2, Grok 4.1, and DeepSeek v3.2 compare as Santa agents
OP here -- I work at Veris and built this. Happy to answer questions about the methodology!
_josh_meyer_··on GPT-5.2, Grok 4.1, and DeepSeek v3.2 compare as Santa agents
SantaBench, a fun benchmark with a serious methodology. The task: play a cheeky Santa agent who researches users online and roasts them based on their social media.
_josh_meyer_··on AbsenceBench: Language models can't tell what's missing
also, not everyone has access. I used (afaik) the only "academic paper --> narrated video" converter, and it just went into beta a couple weeks ago
_josh_meyer_··on AbsenceBench: Language models can't tell what's missing
Wasn't meant to imply the opposite. The video even has a watermark clearly saying it's generated. I genuinely found the video useful, so decided to share.
_josh_meyer_··on What about the Island on ArXiv?
excellent post -- I turned it into a video :) https://supabase.manatee.work/storage/v1/object/public/video...
_josh_meyer_··on MCP in 15min
complete overview of the Model Context Protocol
_josh_meyer_··on Training Code Released for XTTS
Code and a recipe for XTTS_v1.1 GPT encoder training is released under the Mozilla Public License 2.0
_josh_meyer_··on Coqui TTS v0.18
XTTS model release (Text-to-Speech and voice cloning)

# From the release notes:

This model is trained on top of XTTS v1, using output masking. We mask the part of the output that is used as the audio prompt while training and don't compute loss for that segment. This helps us to resolve the hallucination issue that V1 experienced.

- Add Japanese - Resolve the hallucination issue (repeating the audio prompt) - Increased expressivity - Added ne_hifigan that was trained without denoising that brought some EQ and compression profile that might be unwanted for some use-cases

_josh_meyer_··on Voice Chat with Mistral 7B
XTTS + Whisper + Mistral 7B
_josh_meyer_··on XTTS: New Generative model for Voice (weights released on HF)
Coqui releases model weights for XTTS generative Voice model. Demo live on Huggingface
_josh_meyer_··on How to make AI video in 60 seconds
an example pipeline for midjourney / runwayML / coqui to create a video
_josh_meyer_··on Zero Shadow Day
https://en.wikipedia.org/wiki/Zero_shadow_day
_josh_meyer_··on Show HN: AI prompt-to-storyboard videos w/ GPT, Coqui voices, StabilityAI images
Oh this is cool:)
_josh_meyer_··on Coqui Launches Prompt-to-Voice (DALL-E for Voice)
hi there! what kind of better API support do you mean?
_josh_meyer_··on Prompt-to-Voice: create a new voice with Generative AI
Prompt-to-voice creates a new, unique voice given a text prompt. Similar to stable diffusion, but for voices. Then those voices can generate speech via TTS.
_josh_meyer_··on [dead]
Newest release from https://coqui.ai

Direct-able, controllable generative AI Voices

_josh_meyer_··on [dead]
Voice cloning from a segment of Merkel via Coqui.ai
_josh_meyer_··on Sequoia's Map of Generative AI
Map of interesting companies and applications in the Generative AI Landscape, according to Sequoia (Oct. 2022)
_josh_meyer_··on Waitlist for AI Voice Studio
Waitlist for early access to the Coqui Studio
_josh_meyer_··on AI can now sing like Freddy Mercury
imo this kind of tech is useful to supplement the artist, not replace them. You listen to a singer because you know their voice more than anything. Bob Dylan was a pretty bad singer, but I listen to him because of some emotional connection.
_josh_meyer_··on AI can now sing like Freddy Mercury
From the Coqui TTS project <https://github.com/coqui-ai/TTS>
Page 1 of 2Next →