HNHacker News
TopNewBestAskShowJobs

mekpro

589 karma · joined March 24, 2012

submissionscomments
mekpro··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Why they are not even benchmark model against Anthropic or anybody ?
mekpro··on Asus Bike Booster
It’s interesting in terms of the economics of development.

This product takes advantage of advancements in lithium-ion battery technology driven by phones and cars, benefiting from improvements in both performance and cost at scale without requiring extensive additional research. Looking forward to trying it.

mekpro··on Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models
why so much hate on this ? Meta releases open Model and it is competitive and confirms the direction, what could we ask more !!
mekpro··on Previewing GPT‑5.6 Sol: a next-generation model
codex spark is not large model though, much weaker than standard model.
mekpro··on Previewing GPT‑5.6 Sol: a next-generation model
We need more coding benchmark score. Not sure that winning terminalbench 2.1 alone is a clear win over Fable/Mythos yet.
mekpro··on Claude Fable 5 by Anthropic, releasing tomorrow
source ?
mekpro··on Meta Keeps Delaying the Release of Its New AI Model to Developers
API server is not hard problem and not make sense for indefinite postpone. I think the more likely explanation is model quality.

Too bad for Meta, and very sad day Llama.

mekpro··on MAI-Code-1-Flash
The technical report is very detailed and would 'reinforcement learning' of future researchers, Thanks Microsoft!
mekpro··on Expanding Project Glasswing
Yes, 300 MW from SpaceX helps a lot, but I think that’s mainly to support Opus demand, which has grown faster than expected. If Mythos is roughly 5× more expensive to serve than Opus, as the pricing suggests, then 300 MW is nowhere near enough to enable large-scale deployment of Mythos.

As an ordinary developer who relies on a $20–$200/month subscription, I feel disappointed by the release of a paper describing a model that I can’t actually use.

mekpro··on Expanding Project Glasswing
It’s clear that Anthropic has run out of the compute capacity needed to serve Mythos publicly.

They’re using security concerns to mask their inability to deliver the model at scale, while still trying to maintain their lead over OpenAI. As a result, they’ve chosen to release it privately under the banner of an “ethical” rollout.

mekpro··on Kimi Claw
Opus is definitely in its own league. I use Kimi/Gemini-cli code regularly to save cost and from my experience, Kimi 2.5 is more solid than Gemini Flash 3.0 for coding. While Gemini Flash 3.0 is generally faster, it usually break the syntax and skip important prompt. Kimi 2.5 can write very good code and can plan very well.
mekpro··on Kimi Released Kimi K2.5, Open-Source Visual SOTA-Agentic Model
Except that, In OpenRouter, Deepseek always maintain in Top 10 Ranking. Although I did not use it personally, i believe that their main advantage over other model is price/performance.
mekpro··on Apple is fighting for TSMC capacity as Nvidia takes center stage
I think the opposite. Having NVIDIA investing in TSMC's bleeding-edge process node should benefit Apple rather than disadvantage.

It means that Apple doesn't have to be sole investor in latest node development which is more harder to justify, especially in the year where smartphone upgrade cycle is slowdown. Having NVIDIA (and AI boom) in the picture should help Apple reduce CAPEX for their semi-conductor investment.

mekpro··on 1300 Still Images from the Animated Films of Hayao Miyazaki's Studio Ghibli (2023)
They are so beautiful that i dont want any of these been stole by AI.
mekpro··on DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning [pdf]
How this improvement translate into real world agentic coding task ?
mekpro··on Open models by OpenAI
i got 70 token/s on m4 max
mekpro··on Open models by OpenAI
try enable flash attention and offload all layer to GPU
mekpro··on Claude Code weekly rate limits
Is this limit will also count together with Claude Chat ?
mekpro··on OpenAI’s Windsurf deal is off, and Windsurf’s CEO is going to Google
you can easily reach 50$ per day. by force switching model to opus /model opus it will continue to use opus eventhough there is a warning about approaching limit.

i found opus is significantly more capable in coding than sonnet, especcially for the task that is poorly defined, thinking mode can fulfill alot of missing detail and you just need to edit a little before let it code.

mekpro··on Gemini CLI
Just refactored 1000 lines of Claude Code generated to 500 lines with Gemini Pro 2.5 ! Very impressed by the overall agentic experience and model performance.
mekpro··on Ask HN: How to learn CUDA to professional level
To professionals in the field, I have a question: what jobs, positions, and companies are in need of CUDA engineers? My current understanding is that while many companies use CUDA's by-products (like PyTorch), direct CUDA development seems less prevalent. I'm therefore seeking to identify more companies and roles that heavily rely on CUDA.
mekpro··on Devstral
it can use tool to explore directory like ls grep out of the box.
mekpro··on Gemma 3 QAT Models: Bringing AI to Consumer GPUs
Gemma 3 is way way better than Llama 4. I think Meta will start to lose its position in LLM mindshare. Another weakness of Llama 4 is its model size that is too large (even though it can run fast with MoE), which greatly limits the applicable users to a small percentage of enthusiasts who have enough GPU VRAM. Meanwhile, Gemma 3 is widely usable across all hardware sizes.
mekpro··on Google Is Winning on Every AI Front
Google is also the only company that has had their own AI hardware that's worked (TPU). This could lead to more cost-effective training + inference and hence better AI.
mekpro··on Google’s two-year frenzy to catch up with OpenAI
Also, they open-model gemma-3 is very competitive for its size and actually beats llama-3 from Meta. Not to mention that OpenAI doesn't offer anything open anymore.
mekpro··on Palantir Drops 10% on Report of Pentagon Slashing Budget
Still a big bubble considered that the price go up 244% compared to 6 months ago.
mekpro··on QwQ: Alibaba's O1-like reasoning LLM
As a quick estimation, the size of q4 quantized model usually be around 60-70% of the model's parameter. You can preciselly check the quantized model size from .gguf files hosted in huggingface.
mekpro··on Ask HN: What's the "best" book you've ever read?
Zero to One
mekpro··on Tell HN: Merry Christmas
Merry Christmas ! Life is hard but HN always felt like home for me.
mekpro··on Apple's game porting toolkit is fantastic. Cyberpunk 2077 at Ultra on an M1 MBP
This video sample use base M1 chip.
Page 1 of 3Next →