HNHacker News
TopNewBestAskShowJobs

BUFU

646 karma · joined October 28, 2024

submissionscomments
BUFU··on Qualcomm acquires Nexa AI, open-sources GenAI runtime for Hexagon NPUs
blog: https://www.qualcomm.com/developer/blog/2026/06/geniex-devel...
BUFU··on ChatGPT terms disallow its use in providing legal and medical advice to others
Thanks for the clarification. I think if they disallow first parties to get medical and legal advice, it will do more harm than good.
BUFU··on Offline card payments should be possible no later than 1 July 2026
Make card payment available to local AI agents. How would this sound?
BUFU··on Qwen3-VL
The open source models are no longer catching up. They are leading now.
BUFU··on How to increase your surface area for luck
This is the best title I've seen in a while.
BUFU··on Meta's Llama 3.1 can recall 42 percent of the first Harry Potter book
Would it be possible that other people posted content of Harry Potter book online and the model developer scrape that information? Would the model developer be at fault in this scenario?
BUFU··on Local AI generates highly realistic dialogue from a transcript
Saw this on reddit. Super impressive: https://www.reddit.com/r/LocalLLaMA/comments/1k4lmil/a_new_t...

Local NotebookLM Audio Overview coming soon?

BUFU··on Tell HN: Announcing tomhow as a public moderator
Welcome!
BUFU··on NexaQuant: Llama.cpp-Compatible Model Compression with 100%+ Accuracy Recovery
Will llama.cpp be the go-to local inference framework for every device?
BUFU··on OmniAudio-2.6B: Fastest Audio Language Model for Edge Deployment
On a 2024 Mac Mini M4 Pro, Qwen2-Audio-7B-Instruct running on Transformers achieves an average decoding speed of 6.38 tokens/second, while OmniAudio-2.6B through Nexa SDK reaches 35.23 tokens/second in FP16 GGUF version and 66 tokens/second in Q4_K_M quantized GGUF version - delivering 5.5x to 10.3x faster performance on consumer hardware.

Blogs for more details: https://nexa.ai/blogs/OmniAudio-2.6B

HuggingFace Repo: https://huggingface.co/NexaAIDev/OmniAudio-2.6B

Run locally: https://huggingface.co/NexaAIDev/OmniAudio-2.6B#how-to-use-o...

Interactive Demo: https://huggingface.co/spaces/NexaAIDev/omni-audio-demo

BUFU··on What happens if we remove 50 percent of Llama?
This is a crazy thought lol
BUFU··on What happens if we remove 50 percent of Llama?
I believe it definitely does. The inference cost will be much cheaper.
BUFU··on Allen AI released Tülu 3 Models: Open post language model post-training
Hugging Face Repo: https://huggingface.co/collections/allenai/tulu-3-models-673...

Competive with Claude 3.5 haiku, beats all major open models like Llama 3.1 70B, Qwen 2.5 (except MATH) and Nemotron

All their recipe - code, datasets and model checkpoints are public and out in open!

BUFU··on Nvidia presents Llama-Mesh: Generating 3D Mesh with Llama 3.1 8B
From Reddit: https://www.reddit.com/r/LocalLLaMA/comments/1gsohas/nvidia_...

Links

Project Page: https://research.nvidia.com/labs/toronto-ai/LLaMA-Mesh/

HuggingFace Paper: https://huggingface.co/papers/2411.09595

GitHub: https://github.com/nv-tlabs/LLaMA-Mesh