HNHacker News
TopNewBestAskShowJobs

gps372

100 karma · joined January 13, 2017

submissionscomments
gps372··on 9 Ads per Minute: FIFA Cup 26 – "the price of the beautiful game"
This might be already happening. Competitor's market share (without details) is common to share. Scintilla onboarded by Walmart earlier this year might already be giving instore AI capabilities.
gps372··on 9 Ads per Minute: FIFA Cup 26 – "the price of the beautiful game"
True, in-store planogram compliance and floor-plan compliance have got in focus again. It never lost the focus to be honest, just that now there is an expectation to do it while ensuring nothing fall through the cracks.

But there is a much more stringent privacy requirement here since customers walking in those stores haven't really signed up for this analysis. Hence, difficult to just setup the camera and take continuous data for analysis.

gps372··on 9 Ads per Minute: FIFA Cup 26 – "the price of the beautiful game"
Plain and boring computer vision would have still required a team who would identify frames to sample before processing, and then annotate, generalize and classify the data.

Not an expert on this but I don't think same kind of manual effort is required anymore.

gps372··on 9 Ads per Minute: FIFA Cup 26 – "the price of the beautiful game"
Most comments may have ignored this so far, but this is an interesting AI based research.

At the end of the article is this excerpt

>> Results from this research were generated using an AI computer vision model developed by the research team on Bristol’s Isambard-AI, the UK’s most powerful AI supercomputer. Isambard-AI powered the analysis of 172.6 hours of live match footage from all 104 games played between 11 June and 19 July 2026.

So, basically now merchandizers, auditors, compliance monitors, etc. have a tool using which they can quantify whether broadcaster complied with contractual requirements of showing their ads/merch/logos as per contract. And if they missed, then they can quantify the delta - to sue the broadcaster for the balance as well.

Also, it can enable many countries, where certain content is not allowed, then this tool can simply give them data points without having to employ someone, or have someone review the output of the AI tooling.

gps372··on Bend 2 and the Vibe-Coding Trap
Hmm, in your example, temperature is fully measurable but there might be a slight challenge in measuring if "If we are at home". Since for this you need to define more variants with measurements from motion sensors, pressure or weight sensors on beds or sofas, if a device has got connected to wifi, etc.

By the time your first round of beta testing is over, you may have quite a handful of such axioms and variants, which have been humanly validated!

Have you done any such experiment in this space?

gps372··on An empirical study of harness design for coding agents
Haven't gone through full PDF as its very detailed, few things have resonated with me so far.

Basically if a Car A is performing better (be it speed, milage or in general sense) than Car B, then it is not necessarily because its engine. It could be because of better tires, better gearbox, lighter body, better usability of features, etc.

You can implement an AI feature (like AI for BI) in different ways even with the same model - via ReAct-loop, or plan-and-execute, or hybrid. You can make it stateless, stateful, RAG-based, etc. depending upon whether you want to prioritize result accuracy or depth of analysis. You can use LLM to generate either intent (requires lesser reasoning) or the queries itself (requires much more capable model).

Your harness can adapt to the underlying model's native capabilities, or can make up for its absence, e.g. query generation in above example requires your model to have MOE capabilities but intent generation wouldn't.

gps372··on Bend 2 and the Vibe-Coding Trap
Approach itself looked impractical to me for any non-trivial system, like domain centric system of records systems which can have 100s if not 1000s of laws. Though it can be tried as a side parallel thread to see if system is still compliant and following right first principals after a few years from its inception.

I would rather wait to see how it gets adopted, if at all. Anyone aware of early reviews of the adopters of bend 2?

gps372··on How, Exactly, Could A.I. Kill Us?
AI can definitely be a cover for someone to do something nefarious and blame it on the AI.

All that AI needs is access to all the spicy tools like nukes, unsupervised medical diagnosis/treatment, unchecked military decision making, etc. to its React loop so that humans can be completely hands-free and enjoy the fruits of AI's labor.

Maybe we need to mandate parliament (or anything equivalent in other countries) debate before giving any critical tool access to LLM? And then a UN general assembly vote to conclude the process, maybe.

gps372··on OpenSpec – A lightweight and configurable AI spec framework
Thanks for taking time to respond here. Would love to know from your experience the scale of function-points, team size, client-requirement variance, etc. different teams would have worked with and maintained over a period of time via this open-spec.

Please note that I can already see that github repo has 68k+ stars. So popularity is not in question, just the viability and consistency of adoption across different scenarios.

gps372··on OpenSpec – A lightweight and configurable AI spec framework
>> The power is that you do a lot of upfront thinking

This is the best part of this spec, but we have found from our experience that though upfront thinking changes has a lot of merits and adds clarity and alignment upfront, but it changes bit by bit in every meeting and before you know your specs are not aligned with general consensus in the team. If your team is large enough, then it gets very difficult to own the task of constructing alignment between your principal-artifacts and your evolved under-current of understanding.

If you check my submissions (https://news.ycombinator.com/submitted?id=gps372), I have written whole set of articles on the myths of how easy it is keep the understanding consistent.

I would still say that if you are working on a platform and if your engg team size if anything more than 25-30, then this spec must be adopted from top-down and not bottoms up. Bottom level engineers usually don't have the level of consistent exposures (as and when they socialize and evangelize their platform) which top level engineers have.

gps372··on OpenSpec – A lightweight and configurable AI spec framework
Looks like this will be a hard sell for many orgs who are already struggling with explosion of artifacts on JIRA, sharepoint, github, etc. Also, most of them have somewhat settled on some ways (in past 6-8 months) to produce AI first specs and work with them.

Also, this looks like something which leadership level folks need to adopt first and then somehow it needs to trickle down to PI planning and sprint planning. Would like to hear someone's experience on how this has got adopted in their org.

gps372··on Learning Programming in an Age of LLMs
Fully agree. With LLM being able to solve every problem, getting deep into a problem all by yourself becomes a passion side project. Now might be a real test of how much you love programming.

Your enterprise wants the work done, done fast and reliably. Your productivity goals have increased, just like invention of motors would increased goals of carriers who were earlier doing their job via more manual efforts like pedaling. But still people love cycling, but they largely "don't have to" rely on it to do their job.

Similarly, now you simply don't have a dependency to love programming to increase your productivity.

gps372··on Charts built for Chat
Superset comes with a built-in MCP server now https://superset.apache.org/admin-docs/configuration/mcp-ser..., which you need to expose as another container. Any agent with proper authentication (JWT mostly) will be able to create and manage dashboards via natural language.
gps372··on Navier-Stokes Announcement
If it works (something they need to be convinced about), if it accelerate mathematics and solve complex problems for humanity, then why not?

Aren't they already using computers, mobiles, calculators, etc. already?

gps372··on Navier-Stokes Announcement
Agree! 'Problems' are getting solved and this needs to be celebrated. Wondering how this will discourage mathematicians at all, since now they have another tool to accelerate their research. Nothing is stopping them from using 'new technologies' or sticking a gun to their head to use the 'new technologies' either.
gps372··on Astra for Coding: Why Are We Doing This Again?
>> No matter how precisely you specify your epic, if the model will find something that contradicts your knowledge/intent, there are good chances it'll get confused and make subtle errors, and you won't realize until much later.

True! hence the need for someone to review the final spec output and own it as their own output. I have also found LLM to be better at debugging and solving 'a' specific problem, which I believe is due to output's surface area to be reviewed is lesser in comparison.

gps372··on Astra for Coding: Why Are We Doing This Again?
I mean then it might be difficult to get stakeholder's alignment on specs written in typescript though.
gps372··on Astra for Coding: Why Are We Doing This Again?
thanks for sharing, going through your blog about cognitive debt.
gps372··on Astra for Coding: Why Are We Doing This Again?
I have realized that it's like giving a task to a brilliant coder who has just joined the org and is more excited and eager than usual. Hence the responsibility falls squarely on you to set scope constraints while ensuring only to-the-point features are developed.
gps372··on Astra for Coding: Why Are We Doing This Again?
Actually this was a real incident around the end of Feb this year when we had just started experimenting with spec driven development (SDD).
gps372··on Astra for Coding: Why Are We Doing This Again?
Early lesson I learned from AI engineering was - there is no substitute to giving a groomed epic to an agent. Instead of simply saying 'implement themes in my product' you need to be specific, in fact more specific than usual. You need to say exactly what is in scope and what's not, even down to a buttons, events and layouts.

You can groom the epic with the help of AI, but final review must be done by someone who can take ownership of the specs and hence is responsible if something has fallen through the cracks. AI's response will be limited by the output tokens of that specific agent, and there will no repercussions for AI even if it accepts its mistakes.

gps372··on Scientists observe Einstein's gravity in the quantum world
So, basically Gravity is quantum if Gravity is fundamental (not emergent)?

If I am understanding your counterpoint correctly, then this argument can only be concluded satisfactorily if either Graviton is observed (proving is existence) or something more fundamental and deeper is observed (proving that quantum gravity can emerge without graviton).

gps372··on More questions about whether researchers can trust OpenAI with unpublished math
If mathematician was already using OpenAI for research purpose and making progress due to inputs from OpenAI's responses, then I wouldn't put it beyond OpenAI's reach to generate different relevant prompts to make progress by itself. Afterall, Model can keep at it for whatever timeline and keep pursuing all possible combinations it can think try.
gps372··on Scientists observe Einstein's gravity in the quantum world
Maybe I should have been clearer! There is also a possibility for Gravity to be quantum without gravitons https://arxiv.org/abs/gr-qc/0204062 https://en.wikipedia.org/wiki/Induced_gravity https://en.wikipedia.org/wiki/Entropic_gravity

However, for me to claim that quantum Gravity is only emergent without a particle like graviton's mediation, I would need to present evidence. Just like this claim - quantum gravity is only possible from gravitons.

gps372··on Scientists observe Einstein's gravity in the quantum world
>> The only way this statement could be false

That's quite an exotic claim! If Sound and Temperature can emerge without a sound particle or temperature particle, why is it not plausible for gravity to exists without graviton?

Though, I get it that mainstream view from physicists is Graviton is the most 'likely' cause, if the gravity is proven to be quantized. But even they would have the humility to accept that this is a theory yet to be proven and observed!

gps372··on Why the Harness Matters More Than the Model [video]
Harness definitely matters more from safety and reliability point of view, but saying that it matters more than Model itself is a slight exaggeration. In the sense that this 'headline' can lead people to believe that all models are equal and Harness can make up for lack of in-built features of a model.

Internal model must have native features like mixture of experts, memory features, etc. for the harness to use.

gps372··on Programming is Art
>> Who knew?

You must have 'mastered' this art. Unless you are willing to share this mastery and ask for feedback, only you would know!

gps372··on Scientists observe Einstein's gravity in the quantum world
Actually, we do need evidence for premises which are not definitions or basic assumptions (axioms). Graviton is still a hypothetical particle and falls under neither of two categories - Definitions or axioms.
gps372··on US refuses payout for soldiers who died in Iran 'because it isn't a war'
Military Ops, Armed Conflict, Defensive strikes, Global war on terror, counter terrorism ops to name a few. These are the labels given to recent conflicts like Afghanistan, Iraq and Vietnam wars.
gps372··on Programming is Art
Yeah! Thus showing a mirror of the deeper 'you' to yourself. Interesting thought!
Page 1 of 2Next →