HNHacker News
TopNewBestAskShowJobs

BUFU

646 karma · joined October 28, 2024

submissionscomments

Qualcomm acquires Nexa AI, open-sources GenAI runtime for Hexagon NPUs

github.com·5 pts·BUFU·
1

We Ran GPT‑OSS 20B Local on a Phone

nexa.ai·1 pts·BUFU·
0

Qwen3-VL-30B-A3B-Instruct and Thinking

huggingface.co·6 pts·BUFU·
0

New Engine to Run SOTA AI Models on Qualcomm NPU Across Phone, PC, Cars, and IoT

nexa.ai·1 pts·BUFU·
0

First vision language model built off Open AI GPT-OSS

huggingface.co·3 pts·BUFU·
0

First Multimodal AI Model Designed for NPUs

huggingface.co·1 pts·BUFU·
0

Nexa AI Blogs

nexa.ai·1 pts·BUFU·
0

LFM2-VL: Efficient Vision-Language Models

liquid.ai·3 pts·BUFU·
0

Stanford CS336 Language Modeling from Scratch

youtube.com·19 pts·BUFU·
0

Ollama's new app

ollama.com·560 pts·BUFU·
284

You can now connect a directory of apps and tools to Claude with one click

claude.ai·2 pts·BUFU·
1

Ask HN: What tools have you tried to run AI locally on mobile?

2 pts·BUFU·
0

A C++ library to efficiently run Gemma-3N across various platform

github.com·5 pts·BUFU·
0

The Trump-Musk feud has been great for X, which jumped up the App Store charts

techcrunch.com·6 pts·BUFU·
1

How we’re responding to The NYT’s data demands in order to protect user privacy

openai.com·284 pts·BUFU·
324

ChatGPT Deep Research connects cloud apps

twitter.com·1 pts·BUFU·
0

Local AI generates highly realistic dialogue from a transcript

yummy-fir-7a4.notion.site·3 pts·BUFU·
2

Shroud's Spectre Divide and its developer are shutting down

theverge.com·1 pts·BUFU·
0

What Went Wrong with Skype?

theverge.com·4 pts·BUFU·
1

Anthropic's Recommendations to OSTP for the U.S. AI Action Plan

anthropic.com·3 pts·BUFU·
0

Quantized DeepSeek R1 Distill Models with Original Model Accuracy

nexa.ai·2 pts·BUFU·
0

Multimodal Model Quantization Support Through LLM Compressor by Neural Magic

neuralmagic.com·1 pts·BUFU·
0

DeepSeek-R1-Distill-Qwen-1.5B Surpasses GPT-4o in certain benchmarks

huggingface.co·39 pts·BUFU·
17

NexaQuant: Llama.cpp-Compatible Model Compression with 100%+ Accuracy Recovery

nexa.ai·3 pts·BUFU·
1

Meta's new Video Understanding Multimodal Model used Qwen model for training

arxiv.org·7 pts·BUFU·
1

Llama.cpp Now Supports Qwen2-VL (Vision Language Model)

github.com·155 pts·BUFU·
50

OmniAudio-2.6B: Fastest Audio Language Model for Edge Deployment

nexa.ai·2 pts·BUFU·
1

Moondream 0.5B: The Smallest Vision-Language Model

moondream.ai·14 pts·BUFU·
3

ShowUI: One Vision-Language-Action Model for GUI Visual Agent

arxiv.org·2 pts·BUFU·
0

What happens if we remove 50 percent of Llama?

neuralmagic.com·231 pts·BUFU·
132
Page 1 of 2Next →