HNHacker News
TopNewBestAskShowJobs

limoce

1,017 karma · joined April 3, 2021

submissionscomments

The Koala Benchmarks for the Shell

kben.sh·2 pts·limoce·
0

SmallThinker: A Family of Efficient LLMs Natively Trained for Local Deployment

arxiv.org·2 pts·limoce·
0

Step3 Technical Report [pdf]

github.com·1 pts·limoce·
0

FP8 is ~100 tflops faster when the kernel name has "cutlass" in it

twitter.com·243 pts·limoce·
107

Polaris: A Post-training recipe for scaling RL on Advanced Reasoning models

hkunlp.github.io·4 pts·limoce·
1

Overclocking LLM Reasoning: Monitoring and Controlling LLM Thinking Path Lengths

royeisen.github.io·63 pts·limoce·
0

Neutrino: Probing-Based eBPF-Like GPU Kernel Profiling

github.com·2 pts·limoce·
0

Machine Learning Conferences Should Establish "Refutations and Critiques" Track

arxiv.org·3 pts·limoce·
0

SuperGPQA: Scaling LLM Evaluation Across 285 Graduate Disciplines

supergpqa.github.io·2 pts·limoce·
0

SepLLM: Accelerate LLMs by Compressing One Segment into One Separator

sepllm.github.io·39 pts·limoce·
2

Step-Video-T2V: The Practice, Challenges, and Future of Video Foundation Model

arxiv.org·41 pts·limoce·
5

Logic R1: Reproduce DeepSeek R1 Zero on 2K Logic Puzzle Dataset

github.com·1 pts·limoce·
0

Libnginx: Nginx as a Shared Library

github.com·2 pts·limoce·
0

DeepSeek-VL2: Moe Vision-Language Models for Advanced Multimodal Understanding [pdf]

github.com·1 pts·limoce·
0

Fast vectorizable algorithms of binary searching for floating point numbers

github.com·2 pts·limoce·
0

New OpenAI Feature: Predicted Outputs

simonwillison.net·58 pts·limoce·
7

Collaborative Filtering Is Wrong and Here Is Why

link.springer.com·2 pts·limoce·
0

REST: A Plug-and-Play Method for Accelerating LLM Without Additional Training

sites.google.com·1 pts·limoce·
0

Smoke 'em if you got 'em: Hacker gains root access using cigarette lighter

tomshardware.com·2 pts·limoce·
0

O1 Replication Journey: A Strategic Progress Report

github.com·3 pts·limoce·
0

Failures of Gradient-Based Deep Learning (2017) [pdf]

proceedings.mlr.press·1 pts·limoce·
0

Qwen2-VL

huggingface.co·1 pts·limoce·
0

Qwen2-Math

qwenlm.github.io·128 pts·limoce·
38

FlexAttention: The Flexibility of PyTorch with the Performance of FlashAttention

pytorch.org·210 pts·limoce·
24

MiniCPM-v2.6: GPT-4V Level MLLM for Single/Multi Image and Video on Your Phone

github.com·1 pts·limoce·
0

MindSearch: LLM-Based Web Search Engine Similar to Perplexity.ai and SearchGPT

github.com·3 pts·limoce·
0

Turbo Sparse: Achieving LLM SOTA Performance with Minimal Activated Parameters

arxiv.org·11 pts·limoce·
0

PowerInfer-2: Fast Large Language Model Inference on a Smartphone

arxiv.org·1 pts·limoce·
0

Large-scale photonic chiplet Taichi empowers 160TOPS/W AI

science.org·4 pts·limoce·
0

Asterinas: OS kernel written in Rust and providing Linux-compatible ABI

github.com·1 pts·limoce·
0
← PreviousPage 2 of 7Next →