HNHacker News
TopNewBestAskShowJobs

adityashankar

113 karma · joined December 1, 2021

goto https://river.berlin
submissionscomments
adityashankar··on GPT‑Live‑1 in the API
Oh yeah definitely, That being said what helped me more than anything is the flashcards app and speaking German with my roommate regularly.

Nonetheless, in complete honestly I do also have a German tutor who I see once a week for 50 minutes, I am very reliable on completing my work though, the "progressbars" in my flashcards app do keep me motivated.

Learning a language is really hard and takes years, but mentally I am convinced, that if the progressbars in the flashcard app I use reach 100% and also in my German cases app, that I will get closer to speaking perfect German, this keeps me motivated.

adityashankar··on GPT‑Live‑1 in the API
Oh yeah, I built so much stuff to learn German, for example [1] to give me random German texts, force me to read it, and answer it, I created [2] to automatically make flashcards for me and then use with with a flashcards app I regularly use and [3] to help me memorise German cases and word-genders. I love it!

I did think of implementing this conversationally, but tbh it has always been too expensive thus far, I gotta retry with GPT-live-1, I tried it with elevenlabs before but it wasn't live enough and the models were not intelligent enough.

[1] https://river.berlin/projects/german-learning-helper/ [2] https://river.berlin/projects/flashcard-generator/ [3] https://river.berlin/projects/german-cases-trainer/

adityashankar··on OpenRouter is joining Stripe
> For years, OpenRouter has been called “Stripe for LLMs.”

holy crap, save some butter for the bread omg

adityashankar··on I’m leaving OpenAI to build telepathy
This might be naive, but I think the idea is to replace your keyboard
adityashankar··on Ask HN: Who wants to be hired? (August 2026)

  Location:Berlin/Europe
  Remote: Not necessarily, Open to both remote and non-remote opportunities
  Willing to relocate: No
  Technologies:  TypeScript, JavaScript, Python, Vue, React, Svelte, Node, NextJS, Three.js/WebGL, Flask, Django, FastAPI, PostgreSQL, Docker, AWS, GCP, Cloudflare, PyTorch, CUDA, OpenCV, Git, pytest
  Résumé/CV: on https://river.berlin/about-me/
  Email: me@river.berlin
~7 YoE, Senior software developer, I worked previously at Runpod, I was one of the first few people hired there – Looking for MLOps/ML finetuning/Computer Vision/Full stack development related work.

also I am working with huggingface's lerobot right now as a hobby on the side

adityashankar··on Gemini Robotics 2 brings whole body intelligence to robots
I work in this field, and I wrote my bachelor's thesis here, not with humanoids but with VLAs (think chatgpt connected to a robot arm)

It's certainly not there yet for anything practical, there's also certain bits and structures that don't have accurate names during construction, and it is important to keep that in mind - so a robot is unlikely to understand what it means to say "put the left bit of this box onto this right bit" due to ambiguity, a human would understand that

Plus we have no good reliable accuracy testing data in most cases (most tests occur on a few demos, but that isn't a good representation of how must things work), popular benchmarks, such as libero have been saturated, and nearly everything gets 95% there, most companies and researchers have their own benchmarks here.

Plus companies lie alot, and do very dangerous things in thier videos, I.e. these robots should not be standing very close to humans, because of being dangerous.

There are also legitimate concerns of misuse of these robots that need to be accounted for, misuse does not have to be warfare, but can be as simple as confusing it while it is cutting tomatoes with a knife.

Turning doorknob is easy, and fail recovery is also being worked on, but we don't have reliable statistics anywhere on that. The hard part is on practical things, as in when placing bricks or attaching a part during manufacturing it needs to ensure that it is aligning everything correctly....and that's hard, while it is impressive, it is very irresponsible to keep humanoids at home (people are irresponsible when untrained), for example, lawnmowers injure about 6400 people a year...and that is not an everything machine.

Humanoids in general are...not appealing in specific, due to maintainable of joints, complexity, but robot arms in particular, expecially on wheels (check mobile aloha), are likely to be able to do tasks such as clean up in hotels, after a guest had left, or replace some cooks in restaurants (if their work is consistent)

adityashankar··on What does GitHub's security team even do?
I generally believe in hanlons razor (assume stupidity rather than maliciousness), it's likely that github saw the easiest possible solution rather than diving deeper into the cause of the problem to fix it permanently
adityashankar··on OpenAI and Hugging Face address security incident during model evaluation
so openai hacked into huggingface?
adityashankar··on Porting nanochat to a TPU: what carries over from PyTorch, and what breaks
For people that are like me : This entire text is AI generated, i feel weird reading it personally, i guess others may not
adityashankar··on Solving 20 Erdős Problems with 20 Codex Accounts Running in Parallel
tbf I am not able to understand the erdos problem website, as to why it still shows problems as open even if they've (as claimed) claimed to be solved
adityashankar··on Ask HN: What was the last task where only a frontier model could do it?
I was worried about some messup with taxes (my tax advisor messed up here), I managed to get it sorted on time - this was super naive but in -germany when they send you the taxes they mention the "cents" place as well in a very weird way, I assumed that was the entire number and assumed the tax issue I had was 10x larger than the amount it really was (10x and not 100x as my brain is in a place of pressure due to other circumstances and I wasn't thinking clearly)

GPT 5.6 incorrectly stated that I had nothing to do, Fable got the issue correctly and I was able to see that that was indeed the cents place and that I was more worried than I realized, and I managed to get a temporary solution setup (that I verified and I am sure is correct).

Which is to say it managed to relieve me of quite a bit of stress haha

adityashankar··on US seeks cheaper hunter-killer drones after Iran destroys $1B worth of Reapers
How do you even estimate that, human productivity is non linear to worktime and often, employees, when given good benefits are likely to be more productive.

also there's the fact of this having a ROI since people generate economic value in the longer run, and hence, more taxes

adityashankar··on International chess federation sanctions Kramnik
Kramnik has a habit of accusing people of cheating without evidence https://en.wikipedia.org/wiki/Vladimir_Kramnik#Anti-cheating..., he has also done this to many many individuals without evidence (and cheating is taken very seriously), this is in connection to that - he's being penalized for accusing people (repeatedly) without evidence.

Just to be clear – this is an event where he has been doing so repeatedly, accusing people of cheating without evidence in-and-of itself isn't forbidden - since cheating in chess is difficult to verify, it's the manner and fashion to which Kramnik has done this that has led to his suspension

adityashankar··on Ask HN: Who wants to be hired? (July 2026)
Location: Berlin, Germany

Remote: Everything works, I'm open to remote, in-person and hybrid working opportunities

Willing to relocate: No

Technologies :

Pytorch/Tensorflow/Vertex AI/Amazon Bedrock/hf-related libraries, primarily ML-based, but otherwise I'm also familiar with Docker, Python, Pypi, uv, Nodejs, PSQL, SQLite, MS-SQL, GCP, AWS, Datadog, Tinybird, Grafana, Svelte, React, Runpod.

Some Proof : My website is written in Svelte (https://river.berlin) I used to work at Runpod previously, mostly on MLOps, I love training models, or doing the slow work of figuring out vllm/sglang issues. I used to teach python an incredibly long time ago (https://www.youtube.com/@typecodelearn6971) obviously my skills have increased considerably since then. I like deep slow work, things that take a few days to a few hours to solve - That being said I am open to everything, if you want it made, I probably can make it.

Résumé/CV: See it on my website : https://river.berlin

Email : hackernews@river.berlin (or alternatively use the form on my website)

I've 7 years of experience - I started of building services for individuals, and then moved on to ML-related work. I worked at RunPod last, directly working with a number of clients for runpod (including OpenAI, Cursor, Perplexity).

adityashankar··on How is Groq raising more money?
I believe despite quantisation they were still extremely fast, which is still incredibly useful if you don't need high precision/accuracy (which is good enough for many use cases)
adityashankar··on Liquid AI reveals 8B-A1B MoE trained on 38T
This is super interesting, I'm particularly excited for this one as it may allow teams to scale this architecture for VLAs (vision language action models), and having sparser models means more real-time actions on a locally hosted model

demo link for anyone that wants to try this out https://playground.liquid.ai/chat?model=cmppnbgse000004l4bc8...

adityashankar··on Rewrite Bun in Rust has been merged
Curious can you elaborate on this?
adityashankar··on ZAYA1-8B matches DeepSeek-R1 on math with less than 1B active parameters
I used their online api, and asked it to create code for a timer i can copy paste into about:blank to test out (prompt below)

it did it successfully, but it did need a follow up correction prompt, overally pretty impressive for a model with 760M active parameters, but definitely not deepseek-r1 level

that being said, if something with 760M active parameters can be this good then, there's a good chance it is likely that api-based models are likely to get cheaper in the future

Prompt ------

``` can you write me some js code (that i can put in the console for about:blank) which will basically create a timer for for me that i can start, stop, and store current values for (or rather lap)

so i want it to create buttons (start, stop, lap buttons) on the page for me with labels and divs and other elements that accordingly record the current information and display the current information, and can accordingly start, stop and lap :)

the js code that i copy paste automatically creates the html buttons and divs and other elements that can manage the timer and accordingly the timer works with them ```

adityashankar··on Incident with Issues and Webhooks – Resolved
I don't think it's an inferiority complex, negativity sells more and carefully understanding things doesn't sell as much
adityashankar··on Show HN: 1-Bit Bonsai, the First Commercially Viable 1-Bit LLMs
If you look at their whitepaper (https://github.com/PrismML-Eng/Bonsai-demo/blob/main/1-bit-b...) you'll notice that it does have some tradeoffs due to model intelligence being reduced (page 10)

The average of MMLU Redux,MuSR,GSM8K,Human Eval+,IFEval,BFCLv3 for this model is 70.5 compared to 79.3 for Qwen3, that being said the model is also having a 16x smaller size and is 6x faster on a 4090....so it is a tradeoff that is pretty respectable

I'd be interested in fine tuning code here personally

adityashankar··on Show HN: 1-Bit Bonsai, the First Commercially Viable 1-Bit LLMs
The link didn't work for me personally, but that may be a bandwidth issue with me fighting for a connection in the EU
adityashankar··on Show HN: 1-Bit Bonsai, the First Commercially Viable 1-Bit LLMs
here's the google colab link, https://colab.research.google.com/drive/1EzyAaQ2nwDv_1X0jaC5... since the ngrok like likely got ddosed by the number of individuals coming along
adityashankar··on The path to ubiquitous AI (17k tokens/sec)
This depends on how much better the models will get from now in, if Claude Opus 4.6 was transformed into one of these chips and ran at a hypothetical 17k tokens/second, I'm sure that would be astounding, this depends on how much better claude Opus 5 would be compared to the current generation
adityashankar··on Fei-Fei Li's World Labs raised $1B from A16Z, Nvidia to advance its world models
Robotics benefits nearly every industry dependent on robotics, being able to expand your market base to countries where individuals have lower incomes obviously provides a boost to companies that wish to sell more product.

It also allows for companies to reclaim their supply chains within the country of manufacture (for high income countries)

adityashankar··on Tell HN: Another round of Zendesk email spam
I just got 50 emails lol, this really sucks, phew glad i am not alone
adityashankar··on The Startup Graveyard
The filters don't actually work, unless I'm misunderstanding them >.<, can you fix them please?
adityashankar··on Zebra-Llama – Towards efficient hybrid models
yes!, thanks for the link!
adityashankar··on Zebra-Llama – Towards efficient hybrid models
yup that's what I meant!, Jevon's paradox applies to resource usage in general and not towards a specific companies dominance

if computational efficiency goes up (thanks for the correction), and CPU inference becomes viable for most practical applications, GPUs (or accelerators) themselves may be unnecessary for most practical functions

adityashankar··on Zebra-Llama – Towards efficient hybrid models
Due to perverse incentives and the historical nature of models over-claiming accuracy, it's very hard to believe anything until it is open source and can be tested out

that being said, I do very much believe that computational efficiency of models is going to go up [correction] drastically over the coming months, which does pose interesting questions over nvidia's throne

*previously miswrote and said computational efficiency will go down

adityashankar··on Cloudflare was down
it's fine now...I believe
Page 1 of 2Next →