275 karma · joined February 8, 2025
Second, sparse attention is an old area of active research. Offloaded N-gram tables are the next big open weight technological leap.
Second is what my startup specializes in: the offensive part of security needs to become widely available. Asking if your system is patched tells you nothing about whether someone on the open internet with a model can break in. We have a few articles around SIEM evasion + new defensive methods and it's not pretty. An LLM in a good harness now are smart enough that you can get the equivalent of $50k human pentest from a few years ago for a few hundred dollars now.
In hindsight the emergent swarm obviously came from several capabilities built into the models, such as work delegation (subagents) collaboration (GPT Pro-like ensamble), exhaustive exploration (long running agents) hacking (the specific goal of that RL).
Chinese model providers are in a bind because charging a reasonable margin will put them in competition with all the neoclouds that host their model weights without the development cost. Their economics are far more challenging than OpenAI/Anthropic/Google, who already have thriving high revenue business and mountains of compute online or coming online.