HNHacker News
TopNewBestAskShowJobs

mvoodarla

410 karma · joined September 8, 2019

https://mokshith.xyz/

mokshith@sievedata.com

https://www.sievedata.com/

submissionscomments

Building a robust ball tracking system for sports with SAM 2

sievedata.com·1 pts·mvoodarla·
0

Veo 2: Our video generation model

deepmind.google·587 pts·mvoodarla·
327

Guide to pure-audio and audiovisual speaker recognition techniques

sievedata.com·1 pts·mvoodarla·
0

SieveSync: Realistic, zero-shot lipsync pipeline using MuseTalk and LivePortrait

github.com·1 pts·mvoodarla·
0

SieveSync: High-quality, zero-shot lipsync built with MuseTalk and LivePortrait

sievedata.com·4 pts·mvoodarla·
0

Running Meta's SAM2 2x faster

sievedata.com·3 pts·mvoodarla·
0

API to automate social video editing with AI

twitter.com·1 pts·mvoodarla·
0

Finding highlights in long-form video automatically with custom search terms

sievedata.com·1 pts·mvoodarla·
0

Describe Beta: The most descriptive audiovisual summaries for videos

github.com·2 pts·mvoodarla·
1

AI-generated sound effects for stock videos using CogVLM and AudioLDM

sievedata.com·1 pts·mvoodarla·
0

AI active speaker detection on video with a 90% speedup

sievedata.com·3 pts·mvoodarla·
0

Masked Audio Generation Using a Single Non-Autoregressive Transformer

pages.cs.huji.ac.il·1 pts·mvoodarla·
0

Masked Audio Generation Using a Single Non-Autoregressive Transformer

arxiv.org·1 pts·mvoodarla·
0

The most cost-effective audio transcription API

sievedata.com·3 pts·mvoodarla·
0

Audiobox Demo: Where anyone can make a sound with an idea

audiobox.metademolab.com·3 pts·mvoodarla·
0

Audiobox: Generating audio from voice and natural language prompts

ai.meta.com·3 pts·mvoodarla·
1

Improving on open-source for fast, high-quality AI lipsyncing

sievedata.com·2 pts·mvoodarla·
0

Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning

twitter.com·3 pts·mvoodarla·
0

Show HN: State of the art audio enhance (open source AudioSR and DeepFilterNet)

sievedata.com·8 pts·mvoodarla·
1

DA-Clip: Controlling Vision-Language Models for Universal Image Restoration

twitter.com·1 pts·mvoodarla·
0

Voicebox: The first generative AI model for speech to generalize across tasks

ai.facebook.com·7 pts·mvoodarla·
2

DINOv2: State-of-the-art computer vision models with self-supervised learning

dinov2.metademolab.com·176 pts·mvoodarla·
16

Show HN: TrackObject – Drag-and-drop computer vision object tracking

trackobject.xyz·23 pts·mvoodarla·
3

Pix2Seq: A New Language Interface for Object Detection

ai.googleblog.com·1 pts·mvoodarla·
0

Ask HN: Building vision systems without knowing ML?

1 pts·mvoodarla·
0

Launch HN: Sieve (YC W22) – Pluggable APIs for Video Search

sievedata.com·71 pts·mvoodarla·
14

Show HN: Processing 24 hours of video in ten minutes

sievedata.com·77 pts·mvoodarla·
37

Turning petabytes of raw video data into a high-quality ML dataset

medium.com·3 pts·mvoodarla·
2