HNHacker News
TopNewBestAskShowJobs

artemisart

323 karma · joined March 6, 2016

submissionscomments
artemisart··on I resigned from Anthropic today
> First, and least important, consider that self-replicating, solar-powered factories aren't magic; they're algae.

But what is that supposed to mean? Because humanity is not facing existential threat from algae.

artemisart··on Tl;dv: Over 180k meetings left wide open
Interested also, which text-to-speech model do you use? For diarisation Granola uses a chrome extension instead of a bot if that can give you ideas.
artemisart··on Qwen3.8 Max now ranked as the best overall model by agentic index
They didn't run all benchmarks. It's the best in AA agentic index (GDPval-AA v2, ³-Banking) but not coding index (DeepSWE which is missing, Terminal-Bench v2.1 they have 81% vs 90% for Sol, SWE-Atlas-QnA missing).
artemisart··on Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard
Yes same website https://artificialanalysis.ai/models#intelligence-comparison... but they don't have graphs for the individual benchmarks sadly.
artemisart··on Claude Opus 5
No it's a mean of 5 runs.

> We report FrontierCode’s overall score, a composite measure that grades each patch on blocking functional criteria (held-out unit tests) together with weighted code-quality rubric criteria, as mean@5.

They don't explain more in the system card, I guess higher effort levels could loose points on the code quality / scope / style / maintainability stuff?

artemisart··on Kimi K3: Open Frontier Intelligence
Official API doc says only max effort level is currently supported. https://platform.kimi.ai/docs/guide/kimi-k3-quickstart#think...
artemisart··on Claude Fable is relentlessly proactive
bwrap is builtin in claude too, activate with /sandbox command.
artemisart··on MAI-Code-1-Flash
Hill climbing doesn't mean much but absolutely doesn't imply they cheat on benchmarks. They have more details here https://microsoft.ai/news/introducing-mai-thinking-1/ it seems to be "RL on everything".
artemisart··on NSA is using Anthropic's Mythos despite blacklist
Every US intelligence org probably has at least API access, but anything outside of the US? No chance.
artemisart··on OpenZL: An open source format-aware compression framework
I may be misunderstanding the question but that should be just decompressing gzip & compressing with something better like zstd (and saving the gzip options to compress it back), however it won't avoid compressing and decompressing gzip.
artemisart··on GPT-5-Codex
Does refactoring mean moving things around for people? Why don't you use your IDE for this, it already handles fixing imports (or use find-replace) and it's faster and deterministic.
artemisart··on DuckDB NPM packages 1.3.3 and 1.29.2 compromised with malware
Do you know about other security issues? If it's only about curl | sh it really isn't a problem, if the same website showed you a hash to check the file then the hash would be compromised at the same time as the file, and with a package manager you still end up executing code from the author that is free to download and execute anything else. Most package managers don't add security.
artemisart··on DuckDB NPM packages 1.3.3 and 1.29.2 compromised with malware
Why should we expect companies to be able to reuse the correct token if they can't coordinate on using a single domain in the first place?
artemisart··on Mistral raises 1.7B€, partners with ASML
Yes for parakeet, but only comparing benchmark results for canary. Whisper also has severe hallucinations on silence and noise and WhisperX helps a lot, it adds voice activity detection i.e. a model to detect when someone speaks, to filter the input before running whisper. https://github.com/m-bain/whisperX
artemisart··on Mistral raises 1.7B€, partners with ASML
Nvidia parakeet and canary are better and faster, here is a leaderboard: https://huggingface.co/spaces/hf-audio/open_asr_leaderboard
artemisart··on Liquid Glass? That's what your M4 CPU is for
No, you never compute individual pixels because you never need to, and it's always faster to it in bulk (vectorization, memory access...) and so over an area you take the same number of pixels as input (or a little bit more with padding) and the blur will only increase significantly the compute.
artemisart··on What's New with Firefox 142
This seems to be exclusive to Safari, I can't get it to work in Chrome either (and didn't know about the feature before right now, the discoverability is terrible).
artemisart··on Nvidia DGX Spark
I don't understand what's not optimized on 5090. If we're comparing with Apple chips or AMD Strix Halo yes you will have very different hardware + software support, no FP4 etc. but here everything is CUDA, Blackwell vs Blackwell, same FP4 structured sparsity, so I don't get how it would be honest to compare a quantized FP4 model on Spark with an unoptimized FP16 model on a 5090 ?
artemisart··on Nvidia DGX Spark
Ok then just to clarify: you can fit 4x larger models on the Spark vs 5090, not 17x.
artemisart··on Nvidia DGX Spark
That's very true and what's segmenting the market, but I don't understand why you're saying the 5090 supports only 12B model when it can go up to 50-60B (= a bit less than 64B to leave room for inference) as it supports FP4 as well.
artemisart··on YouTube made AI enhancements to videos without warning or permission
The economics don't make sense, each video is stored ~ once (+ replication etc. but let's say O(1)) but viewed n times, so server-side upscaling on the fly is way too costly and currently not good enough client-side.
artemisart··on YouTube made AI enhancements to videos without warning or permission
> Then, once that is perfected, they will offer famous content creators the chance to sell their "image" to other creators, so less popular underpaid creators can record videos and change their appearance to those of famous ones, making each content creator a brand to be sold.

I'm frightened by how realistic this sounds.

artemisart··on Optician Sans – A free font based on historical eye charts and optotypes
Yes I don't understand how they can claim it's optimized for legibility when the base font does the inverse.
artemisart··on Edamagit: Magit for VSCode
Gitless is this fork https://marketplace.visualstudio.com/items?itemName=maattdd.... it's not updated but still works well.
artemisart··on Pyrefly: A new type checker and IDE experience for Python
pyrefly is not tied to vscode? Also please try to be more considerate of people preferences, and pycharm is not strictly better. Remote dev on vscode is very convenient for me, should I go on the Internet saying that pycharm is trash? No
artemisart··on Transformer neural net learns to run Conway's Game of Life just from examples
But it is, as long as the positional embedding are sufficient, i.e. use relative positional embeddings here.
artemisart··on Trump announces 100% tariffs on movies ‘produced in foreign lands’
But do you understand it will harm Hollywood? This is the economic, country scale equivalent of saying "fuck your movies", do you think the answer will be "oh sorry, I'll keep buying yours" or "fuck your movies too"?
artemisart··on Qwen3: Think deeper, act faster
ChatGPT free gets it right without reasoning mode (still explained some steps) https://chatgpt.com/share/6810bc66-5e78-8001-b984-e4f71ee423...
artemisart··on Lossless LLM compression for efficient GPU inference via dynamic-length float
The first sentence of the introduction ends with "we introduce Dynamic-Length Float (DFloat11), a lossless compression framework that reduces LLM size by 30% while preserving outputs that are bit-for-bit identical to the original model" so yes it's lossless.
artemisart··on OpenAI releases image generation in the API
That was the joke.
Page 1 of 3Next →