HNHacker News
TopNewBestAskShowJobs

facu17y

19 karma · joined June 14, 2023

submissionscomments
facu17y··on Safe Superintelligence Inc.
How can they speak of Safety when they are based partly in a colonialist settler entity that is committing a genocide and wanting to exterminate the indigenous population to make room for the Greater Zionist State.

I don't do business with Israeli companies while Israel is engaged in mass Extermination of a human population they treat as dogs.

facu17y··on LoRA from scratch: implementation for LLM finetuning
What's the performance penalty of LoRA?
facu17y··on NY Times copyright suit wants OpenAI to delete all GPT instances
It is legal. Fair use. People have been doing it for ages. Almost every article you've ever read has some fair use of another article, book or news item, etc.
facu17y··on What We Need Instead of "Web Components"
Browser vendors are (or should be) managing the abstractions for their own needs, with developer needs expected to be met by framework/library developers.

Who says web components are meant for use directly by the developer? Maybe they're primarily meant for the browser developers (those who build browser features), not for use directly by web app developers.

facu17y··on Building AI without a neural network
Well, if it's so useless why is it on the HN front page? Are there "PR" companies behind promoting items to the HN front page? I'm sure there are because sometimes an article like this comes up at #3 and everyone says it's got no substance, clickbait, etc
facu17y··on If 95% doesn't count as a vote of no confidence, what number would?
"Four years ago, Altman’s mentor, Y Combinator founder Paul Graham, flew from the United Kingdom to San Francisco to give his protégé the boot, according to three people familiar with the incident, which has not been previously reported."

I guess he had a change of heart about Sam because ... ?

facu17y··on Sam Altman, Greg Brockman and others to join Microsoft
Sam didn't create the breakthroughs behind the current GPT.

He did not create the breakthroughs behind the next GPT.

None of the people that may follow have the same handle on the tech as Ilya. I mean they built up Ilya's image in our mind so much, that he's one of a kind genius (or maybe Musk did that) and now we are to believe that his genius doesn't matter and that Microsoft already knows how to create AGI and that OpenAI is no longer relevant?

Or did I get it wrong?

facu17y··on OpenAI negotiations to reinstate Altman hit snag over board role
You are assuming he wouldn't steal t from OpenAI. He could have a low level employee steal it, and manage to keep it a secret until AGI is born then he takes over the world.
facu17y··on OpenAI's board has fired Sam Altman
What did Sam Altman hide from his board that caused his firing as CEO of OpenAI?

1) That LLMs cannot generalize outside of _patterns_ they pick up during training? (as shown by a recent paper from Google, and as many of us know from our work testing LLMs and working around their short comings)

2) That every time you train a new model, with potentially very high expense, you have no idea what you're going to get. Generally better but also potentially bigger reliability challenges. LLMs are fundamentally unreliable and not stable in any kind of use case besides chat apps, especially when they keep tweaking and updating the model and deprecating old ones. No one can build on shifting sands.

3) The GPT4-Turbo regressed on code generation performance and the 128K window is only usable up to 16K (but for me in use cases more compicated than Q&A over docs, I found that 1.2K is max usable window. That's 100X than he advertised.

4) That he priced GPT4-V at a massive loss to crush the competition

5) That he rushed the GPT Builder product, causing massive drain on resources dedicated to existing customers, and having to halt sign ups, even with a $29B investment riding on the grwoth of the user base. Any one of the above or none of the above.

No one knows... but the board.. .and Microsoft who has 49% control of the board.

facu17y··on Pix2tex: Using a ViT to convert images of equations into LaTeX code
Repo has been deleted? I get a 404. I did see it earlier on.

I fed the equation image (screenshot at the right frame from their gif then cropped) into ChatGPT (GPT4-V) and it correctly deciphered the equation and gave the correct LaText code.

Why was the repo removed?

facu17y··on Establishment of the U.S. Artificial Intelligence Safety Institute
"Despite the increasing complexity and capabilities of machine learning models, they still lack what is commonly understood as "agency." They don't have desires, intentions, or the ability to form goals. They operate under a fixed set of rules or algorithms and don't "want" anything.

Even in feedback loop systems where a model might "learn" from the outcomes of its actions, this learning is typically constrained by the objectives set by human operators. The model itself doesn't have the ability to decide what it wants to learn or how it wants to act; it's merely optimizing for a function that was determined by its creators.

Furthermore, any tendency to "meander and drift outside the scope of their original objective" would generally be considered a bug rather than a feature indicative of agency. Such behavior usually implies that the system is not performing as intended and needs to be corrected or constrained.

In summary, while machine learning models are becoming increasingly sophisticated and capable, they do not possess agency in the way living organisms do. Their actions are a result of algorithms and programming, not independent thought or desire. As a result, questions about their "autonomy" are often less about the models themselves developing agency and more about the ethical and practical implications of the tasks we delegate to them."

The above is from the horse's mouth (ChatGPT4)

My commentary:

We have yet to achieve the kind of agency a jelly fish has, which operates with a nervous system comprised of roughly 10K neurons (vs 100B in humans) and no such thing as a brain. We have not yet been able to replicate the Agency present in a simple nervous system.

I would say even an Amoeba has more agency than a $1B+ OpenAI model since the Amoeba can feed itself and grow in numbers far more successfully and sustainably in the wild with all the unpredictability in its environment than an OpenAI based AI Agent, which ends up stuck in loops or derailed.

What is my point?

We're jumping the gun with these regulations. That's all I'm saying. Not that we should not keep an eye and have a healthy amount of concern and make sure we're on top of it, but we are clearly jumping the gun since we the AI agents so far are unable to compete with a jelly fish in open-ended survival mode (not to be confused with Minecraft survival mode) due to the AI's lack of agency (as a unitary agent and as a collective).

facu17y··on Phind Model beats GPT-4 at coding, with GPT-3.5 speed and 16k context
I meant that saying "something is inconsistent sometimes" is weird because inconsistency implies "sometimes"
facu17y··on Phind Model beats GPT-4 at coding, with GPT-3.5 speed and 16k context
"We do have issues with consistency sometimes" That's a strange statement. Having issues with consistency means that sometimes the output is wrong. What does it mean to have issues with consistency sometimes ? You're either consistent or you're not.
facu17y··on Progress on No-GIL CPython
Mojo was mature enough for a person I know in the community who ported their Python port of llama2 to it. Also, others pointed out other languages they would rather use.

The rationalization in your response obfuscate the real reason you were triggered to down vote, which is that you are too emotionally vested in Python and afraid to try a better alternative.

facu17y··on Progress on No-GIL CPython
This downvoting to -3 is illustrative of how HN down votes are about territorial warfare, ego, etc. Has nothing to do with logic. You don't like to port your python code to nogil python? I'm gonna downvote you. WTF is wrong with you people.
facu17y··on Progress on No-GIL CPython
I'd rather port my Python code to Mojo and get, multi-threading, SIMD and other speedups
facu17y··on PaLI-3 Vision Language Models
no github?
facu17y··on Microsoft is reportedly losing lots of money per user on GitHub Copilot
Down-voted because...?
facu17y··on Microsoft is reportedly losing lots of money per user on GitHub Copilot
What does that say about OpenAI losing money on ChatGPT Plus esp now with multi-modal which I can only assume is very expensive.
facu17y··on TinyML and Efficient Deep Learning Computing
seems that the course doesn't include Google's highly efficiemt "step by step distillation" method, which was discussed on HN yesterday
facu17y··on A clever “perpetual motion” device [video]
I had the Arabic translated edition of that book! same fond memories!
facu17y··on Fine-tune your own Llama 2 to replace GPT-3.5/4
"to replace GPT-3.5/4"

Very inflated statement when it comes to GPT4 since it is a MoE model with 8 separate models each an expert in one area, and you can't replace all 8 models with one model trained for $19.

I call BS on this claim. Maybe it matches GPT4 in the narrow domain you fine-tune it for, and if that can be done for $19 then for $19*8 you can take OpenAI out of business. That doesn't add up.

facu17y··on Asking 60 LLMs a set of 20 questions
Every now and then GPT4 outputs a wrong answer. It's impossible to build a reliable product on top of GPT4 that is not a simple chat bot.
facu17y··on Asking 60 LLMs a set of 20 questions
It might be trained on this question or a variant of it.
facu17y··on Following pushback, Zoom says it won't use customer data to train AI models
Except Signal's founder probably has/had a connection with the NSA. All security is for making it hard for the common attacker, and hostile countries. The NSA, most likely, has social engineered its way into every stack and every important org.
facu17y··on FedNow Is Live
I think FedNow can serve as the inter-bank settlement rails for CBDC, or the wholesale side of CBDC
facu17y··on Llama 2
If we have the budget for pre-training an LLM the architecture itself is a commodity, so what does llama2 add here?

It's all the pre-training that we look to bigCo to do which can cost millions of dollars for the biggest models.

Llama2 has too small of a window for this long of a wait, which suggests that http://Meta.AI team doesn't really have much of a budget as a larger context would be much more costly.

The whole point of a base LLM is the money spent pre-training it.

But it performs badly out of the gate on coding, which is what I'm hearing, then maybe fine-tuning with process/curriculum supervision would help, but that's about it. .

Better? yes. Revolutionary? Nope.

facu17y··on Experiencing decreased performance with ChatGPT-4
the base model may have been but not necessarily the RLHF fine-tuned layers they might have added or the shortcut they're taking during inference due to such fine tuning (or for perf optimization unrelated to fine tuning.)
facu17y··on Douglas Hofstadter changes his mind on Deep Learning and AI risk
if you have to resort to archaic writing to prove a point about the latest and most advanced piece of technology, ... you're practicing religion, not science.
facu17y··on Douglas Hofstadter changes his mind on Deep Learning and AI risk
Show me how I can kill you with my LLM, or my GAN or Diffusion Model.

I dare you!

Page 1 of 2Next →