HNHacker News
TopNewBestAskShowJobs

bertili

478 karma · joined August 15, 2024

submissionscomments
bertili··on Native apps written in TypeScript and CSS
Not apparent at first, but this is a TypeScript to c++ compiler at the core (https://github.com/geastack/compiler), with bindings for various platforms.
bertili··on Xiaomi MiMo v2.6
They mixed up DeepSeek 4.1 Flash with something else on this page, possibly DeepSeek 4.1 Flash means Gemini 3.8 Flash.
bertili··on How, Exactly, Could A.I. Kill Us?
The evolutionary way: Something that is very good at copying itself, will copy itself and gobble up resources, which are finite. Humans took the natural resources from animals and infinite scrolling took the mind-resources from kids. Now animals and our children are fewer and have difficulties to reproduce.
bertili··on DeepSeek v4.1 Flash
The bigger story is the compute efficiency - its been running at 300t/s the last days.
bertili··on Muse Spark 1.3
DeepSWE scores 75.4 - that's the best score so far. And it's crazy cheap! Google held the top a few hours today with Gemini 3.8 Flash, but now second to Spark 1.3. All this competition will drive prices down!
bertili··on Gemini 3.8 Flash and 3.8 Flash Cyber
A fifth of the cost of Opus 5! Google is certainly pushing the completion with this.
bertili··on GLM-5.3-Flash
0 days later: Qwen 3.8 Flash Next: Let's cut GLM 5.3 Flash parmeters in half and active parameters to a third!

Chinese models had 94% reduction in parameters (from 2.8T/104B to 180B/6B) in 6 weeks, while staying close to the same quality.

bertili··on GLM-5.3-Flash
The point is open AI. "Open" as in open weights, open research, open future.
bertili··on GLM-5.3-Flash
This is going so fast! What a time to be on hackernews:

July 16th: The "Kimi K3 moment" - China has caught up to Opus!

4 weeks later: GLM 5.3 - Same performance, but cut the amount of parameters and cost to a third!

12 days later: GLM 5.3 Flash - Almost GLM5.3 performance but cut the parameters in half, cut prices to a fifth and serving on Chinese chips!

bertili··on Qwen3.8 27B scores 52 on Artificial Analysis
I can't shake this the existential feeling that this compact series of 27G bytes represent something profound and universal.
bertili··on Qwen3.8 27B scores 52 on Artificial Analysis
And more context:

Same score as the latest DeepSeek Flash 0731 which has 284B parameters! (13B active)

Its also the second best Qwen model, much better than Qwen 3.7 Max, but significantly below Qwen 3.8 Max.

bertili··on Qwen 3.8 27B
Wow. Speed improved as well. 200t/s on a RTX 5090!

https://x.com/sgl_project/status/2088281320422322413

bertili··on GLM-5.3: Frontier coding with emergent cyber capabilities
This will be roughly on pair with Kimi K3, but using a third of its parameters.

Just 4 weeks ago the "Kimi K3 moment" was seen as a threat to Closed AI and in less than a month Z.ai have cut the parameter/RAM barrier to a third.

Congratulation to Z.ai and all the hard working Chinese researchers who are quitely boiling the frog.

bertili··on GLM-5.3: Frontier coding with emergent cyber capabilities
Musk: Open Chinese models will rival Fable 5 in Q1 2027

JieTang (Founder of Z.ai): It won't take that long

https://x.com/i/trending/2067626647050670400?lang=en

bertili··on GLM-5.3: Frontier coding with emergent cyber capabilities
DwarfStar (https://github.com/antirez/ds4) supports GLM 5.2 and DeepSeek. Not only for toying, but for getting work done.
bertili··on Qwen3.8-Max: A New Bar for Coding and Cowork
The 27B have many more active parameters than much bigger models such as DS4Flash, MiniMax etc, which makes it punch above its tiny weight. A great fit for a 5090 in a closet for meat-and-potatoes, kind of work.
bertili··on Kimi-K3 Releases on HuggingFace 7/27
That looks promising! As models become a commodity, this may turn out to be the real AI gold rush.
bertili··on Kimi-K3 on HuggingFace
Is there any (near future) technology that would permit burning this terrabyte into some kind of ROM chip?
bertili··on Qwen 3.8
Wait.. the Qwen Max models have never been open-weight. But it sure sound like that's what they intend now?

"Qwen3.8 is launching and going open-weight soon! With a massive 2.4T parameters..."

bertili··on Codex Micro
AGI is almost here, but first, one more thing... a keyboard controller!
bertili··on The infinite scroll may become endangered if controversial Calif. law passes
Legislators, please require 10 seconds of load screen with a picture of a tree, for every online video. It worked for cigarette packs.
bertili··on GLM-5.2 is the new leading open weights model on Artificial Analysis
This is GLM 5.2 Max. GLM 5.2 High which use less than half[1] the tokens.

[1] https://z.ai/blog/glm-5.2

bertili··on RTX 5080 and RTX 3090 Setup: 80 Tok/s on Qwen 3.6 27B Q8
Qwen 27b is a compute heavy dense model.
bertili··on Orthrus-Qwen3: up to 7.8×tokens/forward on Qwen3, identical output distribution
Does this translate into a similar reduction in compute?

What's the catch?

bertili··on DeepSeek 4 Flash local inference engine for Metal
equals 2 or 3 human brains in power usage. Amazing work!
bertili··on Qwen3.6-35B-A3B: Agentic coding power, now open to all
It's fascinating that a $999 Mac Mini (M4 32GB) with almost similar wattage as a human brain gets us this far.
bertili··on Qwen3.6-35B-A3B: Agentic coding power, now open to all
Is there any source for these claims?
bertili··on Qwen3.6-35B-A3B: Agentic coding power, now open to all
A relief to see the Qwen team still publishing open weights, after the kneecapping [1] and departures of Junyang Lin and others [2]!

[1] https://news.ycombinator.com/item?id=47246746 [2] https://news.ycombinator.com/item?id=47249343

bertili··on Google releases Gemma 4 open models
The timing is interesting as Apple supposedly will distill google models in the upcoming Siri update [1]. So maybe Gemma is a lower bound on what we can expect baked into iPhones.

[1] https://news.ycombinator.com/item?id=47520438

bertili··on Google releases Gemma 4 open models
Qwen: Hold my beer

https://news.ycombinator.com/item?id=47615002

Page 1 of 3Next →