HNHacker News
TopNewBestAskShowJobs

EarlyOom

205 karma · joined March 22, 2022

submissionscomments

Replace OCR with Vision Language Models

github.com·292 pts·EarlyOom·
125

Show HN: Visually parse an entire YouTube video frame by frame

github.com·5 pts·EarlyOom·
0

Ask HN: What are folks using to train/fine-tune Vision Language Models

1 pts·EarlyOom·
0

A Node.js SDK for calling Vision Language Models

github.com·6 pts·EarlyOom·
0

Run structured extraction on documents/images locally with Ollama and Pydantic

github.com·170 pts·EarlyOom·
29

Show HN: Vlm Run, Extract JSON from images, videos and documents in a simple API

vlm.run·2 pts·EarlyOom·
0

Fine-grained Visual Transcription for YouTube videos

vlm-docs.nos.run·9 pts·EarlyOom·
3

"Ok Computer, why are you slow?"

scottloftin.substack.com·2 pts·EarlyOom·
0

Show HN: NOS – A fast, and ergonomic PyTorch inference server

github.com·3 pts·EarlyOom·
0