HNHacker News
TopNewBestAskShowJobs

Oras

2,419 karma · joined January 24, 2014

submissionscomments
Oras··on Dots: Always-on agents
Surprised people are comparing to muse. Meta reputation is terrible within HN audience, so for people to use their AI agent as an example feels like “muse generated” argument. Unless the sentiment has change quite recent and I missed it
Oras··on Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms
My use case is simple classification for job ads. Things like, industry, work settings (remote, hybrid, onsite) and job type (full time, part time .. etc).

I did side by side comparison with Gemini 2.5 Flash Lite, Jev, Jeff

I tried the 0.8B model, completely useless in classification. Qwen Jeff-Qwen3.5-2B was better, but still missed job type.

I suppose with larger model, this could be useful, but would require more ram and will be slower.

Oras··on Revealing the details of how OpenAI agents hacked Hugging Face
So agents made all these chained short URLs that runs code which is pretty clever, but at which point and how they had access to internal HF systems? Were these sandboxes running inside HF production platform?
Oras··on Strands Harness
I remember reading about strands SDK and it looked great in terms how everything is an event that you can extend, so this harness feels quite about right.

However, for this kind of customisation, Pi is actually quite great. One of the most things I love about Pi is ability to ask it to create an extension and it does it quite well as it’s part of their docs. Also ability to customise the system prompt to avoid the clutter that Claude Code add (around 20k system prompt that mostly had nothing to do with the code).

The demo was showing something I have created for my Pi setup, which is asking me in each new session which skills and MCP I would to enable for the session. This works quite well if you have multiple projects where you don’t need all skills but just a small subset

Oras··on Show HN: Matcha – fresh job postings pulled directly from company career pages
I tried searching jobs in London, showed me jobs from all over the UK. I also tried "Outside IR35" which is a type of contracting in the UK, got lots of irrelevant results.
Oras··on Kev: Tiny Jev-like family of decision models built on top of Qwen3.5
For a while is the keyword. It’s just vibe coders have just discovered the classifiers
Oras··on Ask HN: How do you interview devs in a post-AI world?
Leetcode doesn’t tell you anything apart from someone memorised set of questions and their solution.

For your question, I assume you want people who can solve problems, can explain their thought process and reasons of decisions. I did interview devs recently, and my questions were related to their experience in their CV. Something like, tell me about project X in Y company, what did you do, what did you use, why, what did you learn from it.

My personal opinion is that LLMs are no difference from developers who over engineer and over complicate things. You can get them to do good work with the right steering and right understanding of the whole system within the context of the company.

Oras··on I built non-autoregressive decision models with RL a year ago
I made it clear that it is useful and I can see many people using it including myself. My point is it’s not a breakthrough.
Oras··on I built non-autoregressive decision models with RL a year ago
I wouldn’t say lazy, LLMs are fast to use and much more cost effective especially if you factor the cost and time of training (data preparation, data cleaning, … etc).

It’s hard to justify several months to business when there is something off-shelf ready to use and doesn’t require domain specialists to run.

Oras··on I built non-autoregressive decision models with RL a year ago
I played around with Jev last night and did it for classification tasks that I used Gemini 2.5 flash lite with.

It’s a bit faster and bit cheaper, but this is compared to LLM. The consistency was nice to see, BUT, as someone who trained NLP models prior to LLMs, it’s just BERT with more data. I can see why people would want ready made one shot classifier, and I can see the value of sending multiple classifier in one call, but I wouldn’t call it breakthrough. And I believe many labs will replicate it in no time and might have it as part of their harness.

I see it as a wake up call for the tech community to go back to basics for most tasks instead of relying solely on generic LLMs.

Oras··on Fuck it, make it anyway
There is nothing stopping anyone doing things their way, LLMs or not.

I suppose you can balance between using them if the paycheck job is asking for them, and enjoy your craft outside your 9-5.

I’ve been reading more and more of this tone, and I do wonder what those devs writing these blogs do and where do they work. I’ve been a software engineer for more than 25 years, and the challenge was never the cording, it was always the human communication and understanding the outcome.

Yes LLMs can help none devs create things now, but show me a company that will allow AI generated code by none devs to go to prod.

Oras··on I spent $220 on Google app ads and 60% of the installs were robots
But even reviews are fake now. 2 years ago I launched a service and added trust pilot to collect some reviews to help with SEO, then received many emails promising “verified” reviews for $x per review.

Same for Google reviews.

I’m a software engineer, and the way I find new tools is usually YouTube, GitHub, and HN. Even those channels are manipulated now with bots to inflate numbers to reach more

Oras··on I spent $220 on Google app ads and 60% of the installs were robots
Reading all the comments, it seems quite bleak to advertise. How do indies and SMBs reach audience these days then?
Oras··on iPhone 18 Pro and iPhone 18 Pro Max
Isn’t that done by EXIF like decades ago? Not on phones but DSLR. I would think it’s an easy software thing to add (and probably to remove)
Oras··on DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
LLMs are not consistent
Oras··on How An AI math breakthrough ignited a controversy
That’s the definition of current LLMs. They have trained “borrowed” on whatever humans have documented digitally and physically (books).
Oras··on Mistral raises €3B
> > I think the most of the money would go to purchase hardware,

> So it's not a skill issue?

I meant by offering sovereign cloud/inference, not for training. But even if it was for training, Chinese labs have limited supply of GPUs, look what they've done. So it is a skill issue.

Also to clarify, I didn't mean European engineers' skills, I meant "you get what you pay for" as a company, that's why I hope they start offering better compensation to retain talent.

Oras··on Mistral raises €3B
Chinese models are open, available to distill, and they also publish papers about their research. Being one year behind is a skill issue.

I think the most of the money would go to purchase hardware, But I hope they can start making adequate compensation for AI engineers and researchers to move forward.

Oras··on Google AI Mode shows same products 21.6% more expensive than traditional search
I tried “mens cycling helmet specialized” via normal search and AI search.

On the surface, AI showed results of £45 while normal results showed £39.99 and this particular result didn’t show at all in the AI search, even when clicking more.

However, small fine print was the £39.99 had a £4.99 delivery. Can it be that AI is optimised for full price including shipping?

Oras··on Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?
I like the theories here, we shall see if it’s another DNS issue
Oras··on Gemini 3.8 Flash and 3.8 Flash Cyber
Sums up Google AI products.

I have a weird vibe from all the comments in this thread, they feel like a script rather a real experience.

Oras··on HomeOS – A self-hosted family dashboard for a kitchen touchscreen
I'm a techie, but nothing beats a magnetic whiteboard on the fridge.
Oras··on Google is catching up in coding, finally
That would be interesting to see. The only time Google had a good coding model was 2.5 Pro, which lasted a few weeks then degraded so bad.

A gentle reminder that none of Google models are in top 10 in Terminal Bench, so if this article is true, that would be a leap.

Oras··on True Rate of Unemployment
If the uni time didn’t trigger basic instinct of “researching” then that was a huge waste of time and money
Oras··on Dyson CameraJet Toothbrush
> Watch live cleaning footage as you brush, or use the intraoral camera for teeth inspection.

What a time to be alive!

That said, it’s really impressive to have a camera and all this tech in a toothbrush

Oras··on Ask HN: Who is hiring? (September 2026)
What a weird choice for domain name
Oras··on Show HN: Pause – weekly curated 1:1 coffee matches for London professionals
AI generated page, no links to who created it, what they do, and somehow making it to HN front page.
Oras··on Ask HN: Coding is a solved problem. What is left for experienced engineers?
Skills and agents md files have better guidance than most senior engineers? Oh well, good for you then, especially with the “trust me” phrase
Oras··on Ask HN: Coding is a solved problem. What is left for experienced engineers?
> The important part seems to be knowing what to ask, what to check, and what you are actually trying to build.

1. Understanding user requirements/pain points is very important and many get it wrong.

2. Doing the architecture of the system within the org env/cloud/infra is often wrong by LLMs (for now).

3. Debugging when things go wrong. What to log, how to log it, and how to ensure not logging much or logging sensitive data.

4. Guide junior engineers, so they are not just accepting what LLMs are spiting.

Oras··on Andreessen Horowitz is investing billions into a bleak future
By your logic, hacking is legal due to vulnerabilities
Page 1 of 32Next →