HNHacker News
TopNewBestAskShowJobs

campers

356 karma · joined June 3, 2014

DevOps/SRE lead for TrafficGuard - Ad fraud protection platform with clients such as Disney Streaming, Gojek, BetFred, Bet365, Singtel. https://www.trafficguard.ai

Java/Kotlin/JS/TypeScript/Node.js/Google Cloud developer in Perth, Australia

https://apporchestra.com

daniel.campagnoli@trafficguard.ai https://www.linkedin.com/in/danielcampagnoli/

meet.hn/city/au-Perth

Socials: - linkedin.com/in/danielcampagnoli

Interests: AI/ML, DevOps, Cybersecurity, Outdoor Activities, Open Source

---

submissionscomments
campers··on Updated Google Maps shows destruction of the city of Rafah
https://www.un.org/en/preventgenocide/rwanda/assets/pdf/Guid...

  The legal definition of genocide is precise and includes an element that is often hard to prove: "intent."
  
  Determining whether a situation constitutes genocide is factually and legally complex. It should only be made following a careful and detailed examination of the facts against relevant legislation.
  
  This examination is carried out to establish State responsibility or individual criminal responsibility for the crime of genocide. It must be done by a competent international or national court with jurisdiction to try such cases, after an investigation that meets appropriate due process standards.
Correct. If no such court case has delivered a verdict, then by definition no one can rightly claim genocide has occurred.

As for whats happening, the over represented proportion of male combatant age deaths, the well below average ratio of civilian to combatant deaths in complex urban warfare, the war being in response to an atrocious attack on civilians, stopping for polio vaccines, population numbers remaining stable etc puts a question on the "Only Reasonable Inference" standard.

What would you consider a proportionate response given Hamas's initial charter, the rockets fired over the years, the IRCG commanders statements about a co-ordinated attack from Gaza, southern Lebanon and the West Bank and the abhorrent attack of October 7th?

  Major General Gholam Ali Rashid—commander of the Khatam ol Anbia Central Headquarters, which is the highest Iranian operational command level and responsible for all joint operations— expanded on this concept in a May 2024 interview and demonstrated how Iran is continuing to learn from Hamas. Rashid argued that the Hamas attack into Israel in October 2023 highlighted how effective and valuable ground attacks could be.[iii] Hamas’ attack demonstrated, Rashid said, that the Axis of Resistance could destroy the Israeli state by launching surprise attacks from Lebanon, the Gaza Strip, and the West Bank simultaneously—echoing Salami’s comments. Rashid stated that such an attack would need 10,000 fighters from Lebanon, 10,000 fighters from the Gaza Strip, and 2,000-3,000 from the West Bank. Rashid’s interview is particularly noteworthy given his importance in Iranian military decision-making and planning as Khatam ol Anbia Central Headquarters commander.
https://www.criticalthreats.org/analysis/how-iran-plans-to-d...
campers··on Claude Fable 5.1 and Claude Mythos 5.1
Each time you re-write keep a copy of the before and after with some notes on why. Then with a few good examples of this turn it into a skill to review/fix new UI copy.
campers··on Mike: open-source legal AI
Interested to try it out! Some feedback on the homepage there's nothing above the fold, or directly below that says its a Legal AI platform. I would like a legal AI tool, but I'm not familiar with the space don't know what Harvey or Legora are. It was only the hackernews title "Mike: open-source legal AI" that gave the context.
campers··on Gemini 3 Flash: Frontier intelligence built for speed
I had wondered if they run their inference at high batch sizes to get better throughput to keep their inference costs lower.

They do have a priority tier at double the cost, but haven't seen any benchmarks on how much faster that actually is.

The flex tier was an underrated feature in GPT5, batch pricing with a regular API call. GPT5.1 using flex priority is an amazing price/intelligence tradeoff for non-latency sensitive applications, without needing to extra plumbing of most batch APIs

campers··on Imec's superconducting chips to shrink power usage 100x (2024)
OpenAI and Nvidia's 10GW datacenter agreement and Sam Altmans new blog post wanting to build 1GW of infra a week made me think of this article again and to post it up.

Here's a couple of interesting paragraphs from it.

  Instead of the transistor, the basic element in superconducting logic is the Josephson-junction.
  
  For logic, a Josephson-junction loop without a persistent current indicates a logical 0, while a loop with one single flux quantum’s worth of current represents a logical 1. For memory, two Josephson junction loops are connected together. An SFQ’s worth of persistent current in the left loop is a memory 0, and a current in the right loop is a memory 1.
  
  In classical CMOS-based technology, it is very challenging to stack computational chips on top of each other because of the large amount of power, and therefore heat, that is dissipated within the chips. In superconducting technology, the little power that is dissipated is easily removed by the liquid helium. Logic chips can be directly stacked using advanced 3D integration technologies resulting in shorter and faster connections between the chips, and a smaller footprint.
  
   It is also straightforward to stack multiple boards of 3D superconducting chips on top of each other, leaving only a small space between them. We modeled a stack of 100 such boards, all operating within the same cooling environment and contained in a 20- by 20- by 12-centimeter volume, roughly the size of a shoebox. We calculated that this stack can perform 20 exaflops (in BF16 number format), 20 times the capacity of the largest supercomputer today. What’s more, the system promises to consume only 500 kilowatts of total power. This translates to energy efficiency one hundred times as high as the most efficient supercomputer today.
campers··on Gemini CLI GitHub Actions
I added a key rotator to my AI coder, and asked a couple of friends to make keys for me. That helped code a good chunk of http://typedai.dev when 2.5 Pro came out
campers··on Gemini 2.5 Deep Think
Google actually does provide that service! https://cloud.google.com/vertex-ai/generative-ai/docs/model-...
campers··on Qwen3-Coder: Agentic coding in the world
Looking forward to using this on Cerebras!
campers··on Gemini Diffusion
I've been thinking about adding in an agent to our Codex/Jules like platform which goes through the git history for the main files being changed, extracts the Jira ticket ID's, look through them for additional context, along with the analyzing the changes to other files in commits.
campers··on Thoughts on thinking
There is a huge focus on training the LLMs to reason, that ability will slowly (or not that slowly depending on your timeframe!) but surely improve in the AI models given the gargantuan amount of money and talent being thrown at the problem. To what level we'll have to wait and see.
campers··on A 10x Faster TypeScript
There isn't a TypeScript runtime, it is just a JavaScript/ECMAScript compiler/transpiler with a type checking and language server
campers··on GPT-4.5
The price will come down over time as they apply all the techniques to distill it down to a smaller parameter model. Just like GPT4 pricing came down significantly over time.
campers··on Show HN: Mastra – Open-source JS agent framework, by the developers of Gatsby
I get the same feeing when I first looked at the LangChain documentation when I wanted to first start tinkering with LLM apps.

I built my own TypeScript AI platform https://typedai.dev with an extensive feature list where I've kept iterating on what I find the most ergonomic way to develop, using standard constructs as much as possible. I've coded enough Java streams, RxJS chains, and JavaScript callbacks and Promise chains to know what kind of code I like to read and debug.

I was having a peek at xstate but after I came across https://docs.dbos.dev/ here recently I'm pretty sure that's that path I'll go down for durable execution to keep building everything with a simple programming model.

campers··on Show HN: Mastra – Open-source JS agent framework, by the developers of Gatsby
https://typedai.dev is another full-featured one I've built, with a web UI, multi-user support, code editing agents, CodeAct autonomous agent
campers··on Llama 3.1 405B now runs at 969 tokens/s on Cerebras Inference
https://web.archive.org/web/20230812020202/https://www.youtu...
campers··on FrontierMath: A benchmark for evaluating advanced mathematical reasoning in AI
If an AI achieved 100% in this benchmark it would indicate super-intelligence in the field of mathematics. But depending on what else it could do it may fall short on general intelligence across all domains.
campers··on Cerebras Trains Llama Models to Leap over GPUs
On Google Cloud a server with 8 TPU v5e will do 2175 token/seconds on Llama2 70B.

https://cloud.google.com/blog/products/compute/updates-to-ai...

From https://cloud.google.com/tpu/pricing and https://cloud.google.com/vertex-ai/pricing#prediction-prices (search for ct5lp-hightpu-8t on the page) the cost for that appears to be $11.04/hr which is just under $100k for a year. Or half that on a 3-year commit.

That seems like a better deal than millions for a few CS-3 nodes.

And they've just announced the v6 TPU:

  Compared to TPU v5e, Trillium delivers: 
  Over 4x improvement in training performance 
  Up to 3x increase in inference throughput 
  A 67% increase in energy efficiency
  An impressive 4.7x increase in peak compute performance per chip 
  Double the High Bandwidth Memory (HBM) capacity 
  Double the Interchip Interconnect (ICI) bandwidth 
https://cloud.google.com/blog/products/compute/trillium-sixt...
campers··on Cerebras Inference now 3x faster: Llama3.1-70B breaks 2,100 tokens/s

  The first implementation of inference on the Wafer Scale Engine and utilized only a fraction of its peak bandwidth, compute, and IO capacity. Today’s release is the culmination of numerous software, hardware, and ML improvements we made to our stack to greatly improve the utilization and real-world performance of Cerebras Inference.
 
  We’ve re-written or optimized the most critical kernels such as MatMul, reduce/broadcast, element wise ops, and activations. Wafer IO has been streamlined to run asynchronously from compute. This release also implements speculative decoding, a widely used technique that uses a small model and large model in tandem to generate answers faster.
campers··on Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
A bit like the new Gemini Pro 1.5-002 release.
campers··on Differential Transformer
The tl;dr on high level performance improvements

"The scaling curves indicate that Diff Transformer requires only about 65% of model size or training tokens needed by Transformer to achieve comparable language modeling performance."

"Diff Transformer retains high performance even at reduced bit-widths, ranging from 16 bits to 6 bits. In comparison, Transformer’s accuracy significantly drops with 6-bit quantization. The 4-bit Diff Transformer achieves comparable accuracy as the 6-bit Transformer, and outperforms the 4-bit Transformer by about 25% in accuracy."

campers··on Canvas is a new way to write and code with ChatGPT
Check out https://sophia.dev Its AI tooling I've built on top of Aider for the code editing. I initially built it before Aider added support for running compile and lint commands, as it would often generate changes which wouldn't compile.

I'd added seperate design/implementation agents before that was added to Aider https://aider.chat/2024/09/26/architect.html

The other different is I have a file selection agent and a code review agent, which often has some good fixes/improvements.

I use both, I'll use Aider if its something I feel it will right the first time or I want control over the files in the context, otherwise I'll use the agent in Sophia.

campers··on Canvas is a new way to write and code with ChatGPT
I mainly use CLI tools for AI assistance.

I'll use Continue when a chat is all I want to generate some code/script to copy paste in. When I need to prepare a bigger input I'll use the CLI tool in Sophia (sophia.dev) to generate the response.

I use Aider sometimes, less so lately, although it has caught up with some features in Sophia (which builds on top of Aider), being able to compile, and lint, and separating design from the implementation LLM call. With Aider you have to manually add/drop files from the context, which is good for having precise control over which files are included.

I use the code agent in Sophia to build itself a fair bit. It has its own file selection agent, and also a review agent which helps a lot with fixing issues on the initial generated changes.

campers··on Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more
https://www.etched.com/announcing-etched

I think there's another one but I can't remember the name of it.

Also a bit further out is https://spectrum.ieee.org/superconducting-computer

"Instead of the transistor, the basic element in superconducting logic is the Josephson-junction."

campers··on Launch HN: Deepsilicon (YC S24) – Software and hardware for ternary transformers
Can it generate Doom at runtime?
campers··on From Opium to Saffron, the Ancients Knew a Thing or Two About Drugs
This is a great read https://www.researchgate.net/publication/331622659_Getting_h...

> This article collects evidence from psychopharmacology, scripture, and archeology to explore several preparations for consumption described in the Old Testament: Manna, Showbread, the Holy Ointment, and the Tabernacle Incense. The Ointment and the Incense are herbal preparations used by the priestly caste to facilitate a direct experience of the Israelite God. A wide variety of psychoactive components are found in these preparations, including GABA-receptor agonists and modulators, opioid receptor agonists, and other agents. They are normally broken down by the body’s enzymes, and therefore orally inactive, but the Holy Ointment also contains inhibitors specific to the enzymes in question. The preparations indicate that the ancient Israelites had a profound understanding of synergism, and the way they are consumed and the taboos around them are highly suggestive of their use as psychoactive agents.

campers··on Anthropic Claude 3.5 can create icalendar files, so I did this
Or three or four or five! https://openreview.net/pdf?id=zj7YuTE4t8
campers··on Accident Forgiveness
> Meanwhile: like every public cloud, we provision our own hardware, and we have excess capacity. Your messed-up CI/CD jobs didn’t really cost us anything

This^ Don't feel bad about asking for credits when you accidentally make a costly mistake.

campers··on Accident Forgiveness
Being at the level where you have an account manager should help. Our technical account manager always said if we ever accidentally rack up a bunch of cloud costs we should at always ask if we can get a refund, and as long as its infrequent there's a good chance to get some credits.
campers··on Show HN: Nous – Open-Source Agent Framework with Autonomous, SWE Agents, WebUI
I'm going to make the availability of requestFeedback function a boolean flag, so when running benchmark suites etc it can be disabled. Whether its an assistant or agent by that definition is just really a parameter value.
campers··on Zed AI
That was a part of the reasoning of open sourcing my AI assistant/software dev project. Companies like Google have strict procedures around access to customer data. The same can't always be said about a startup racing to not run out of cash.
Page 1 of 5Next →