HNHacker News
TopNewBestAskShowJobs

aliljet

1,538 karma · joined February 22, 2016

contact me here: pav.gup@gmail.com
submissionscomments
aliljet··on Stripe valued at $159B, 2025 annual letter
The easiest way? Angelist.
aliljet··on Stripe valued at $159B, 2025 annual letter
The path to declaring yourself accredited is uniquely easy. Just say it. The whole space is deeply unregulated and unaudited. What makes it insane is that those middleman are making a small fortune exploiting this loophole protecting large companies from being forced to go public. The number is 2000 private investors. Rest assured, more than 2000 individuals have money in Stripe today. It's a total scam.
aliljet··on Stripe valued at $159B, 2025 annual letter
The public can absolutely participate in this by way of syndication deals. Those syndicates are what's covering up the true extent of ownership and they're essentially charging for access with their fees. It's oddly shady, poorly regulated, and more expensive than just being public, but everyone can ride this ride.
aliljet··on Qwen3.5: Towards Native Multimodal Agents
That's a bit confusing. Do you believe LLMs coming out of non-chinese labs are censoring information about Israel and/or Palestine? Can you provide examples?
aliljet··on AWS Adds support for nested virtualization
I wonder if this will extend SEV-SNP and TDX to the child VMs?
aliljet··on MiniMax M2.5 released: 80.2% in SWE-bench Verified
I wonder if these are starting to get reasonable enough to use locally?
aliljet··on Gemini 3 Deep Think
The problem here is that it looks like this is released with almost no real access. How are people using this without submitting to a $250/mo subscription?
aliljet··on Warcraft III Peon Voice Notifications for Claude Code
What I really want is for the peon voice to be replicated and for custom things to be in that voice. Or even better, the starcraft battlecruiser guy's voice!
aliljet··on Hologram v0.7.0: Milestone release for Elixir-to-JavaScript porting initiative
Am I just a cynic, or do any of the LLMs deserve love too?
aliljet··on GLM-OCR – A multimodal OCR model for complex document understanding
This is precisely the real question. If you're exceeding human transcription, you may be generally pretty good. The question is what happens when you tell a human to become surgical about some part of the document, how then does the comparison change..
aliljet··on GLM-OCR – A multimodal OCR model for complex document understanding
All of healthcare is crying. Trust me.
aliljet··on GLM-OCR – A multimodal OCR model for complex document understanding
This is actually the thing I really desperately need. I'm routinely analyzing contracts that were faxed to me, scanned with monstrously poor resolution, wet signed, all kinds of shit. The big LLM providers choke on this raw input and I burn up the entire context window for 30 pages of text. Understandable evals of the quality of these OCR systems (which are moving wicked fast) would be helpful...

And here's the kicker. I can't afford mistakes. Missing a single character or misinterpreting it could be catastrophic. 4 units vacant? 10 days to respond? Signature missing? Incredibly critical things. I can't find an eval that gives me confidence around this.

aliljet··on Level S4 solar radiation event
I wonder if we're going to see an aurora over Seattle tonight?
aliljet··on I'm Peter Roberts, immigration attorney who does work for YC and startups. AMA
Is there clarity right now around foreign students attempting to obtain h1bs in the future?
aliljet··on Show HN: I built a dashboard to compare mortgage rates across 120 credit unions
This is absolutely fantastic. I wish this included commercial loans like DSCRs...
aliljet··on Meta Segment Anything Model 3
I wonder how effective this is medical scenarios? Segmenting organs and tumors in cat scans or MRIs?
aliljet··on Claude Opus 4.5
The real question I have after seeing the usage rug being pulled is what this costs and how usable this ACTUALLY is with a Claude Max 20x subscription. In practice, Opus is basically unusable by anyone paying enterprise-prices. And the modification of "usage" quotas has made the platform fundamentally unstable, and honestly, it left me personally feeling like I was cheated by Anthropic...
aliljet··on Adversarial poetry as a universal single-turn jailbreak mechanism in LLMs
This is great, but I was hoping to read a bunch of hilarious poetry. Where is the actual poetry?!
aliljet··on Gemini 3
Thus far, this is one of the best objective evaluations of real world software engineering...
aliljet··on Gemini 3
Understanding precisely why Gemini 3 isn't front of the pack on SWE Bench is really what I was hoping to understand here. Especially for a blog post targeted at software developers...
aliljet··on Gemini 3
This is the heroic move everyone is waiting for. Do you know how this will be priced?
aliljet··on Gemini 3
When will this be available in the cli?
aliljet··on Gemini 3 Pro Model Card [pdf]
What's wild here is that among every single score they've absolutely killed, somehow, Anthropic and Claude Sonnet 4.5 have won a single victory in the fight: SWE Bench Verified and only by a singular point.

I already enjoy Gemini 2.5 pro for planning and if Gemini 3 is priced similarly, I'll be incredibly happy to ditch the painfully pricey Claude max subscription. To be fair, I've already got an extremely sour taste in my mouth from the last Anthropic bait and switch on pricing and usage, so happy to see Google take the crown here.

aliljet··on GPT-5.1: A smarter, more conversational ChatGPT
What we really desperately need is more context pruning from these LLMs. The ability to pull irrelevant parts of the context window as a task is brought into focus.
aliljet··on Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
How does one effectively use something like this locally with consumer-grade hardware?
aliljet··on Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
The system is working! :)
aliljet··on Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
Where is our guy @simonw on this..
aliljet··on Tongyi DeepResearch – open-source 30B MoE Model that rivals OpenAI DeepResearch
oh my god. 128 gb of RAM! way too late to repair this thread, but most people caught this.
aliljet··on Tongyi DeepResearch – open-source 30B MoE Model that rivals OpenAI DeepResearch
Sunday morning, and I find myself wondering how the engineering tinkerer is supposed to best self-host these models? I'd love to load this up on the old 2080ti with 128gb of vram and play, even slowly. I'm curious what the current recommendation on that path looks like.

Constraints are the fun part here. I know this isn't the 8x Blackwell Lamborghini, that's the point. :)

aliljet··on Attention lapses due to sleep deprivation due to flushing fluid from brain
What have you done when your toddler wakes up at random hours during the night to interrupt your sleep and come and play? That's what has truly obliterated our sleep. Everything else was a passing fad that was minimally painful at best..
← PreviousPage 3 of 9Next →