HNHacker News
TopNewBestAskShowJobs

robertnishihara

464 karma · joined October 16, 2017

submissionscomments

High Performance Distributed Inference with Ray Serve LLM

anyscale.com·3 pts·robertnishihara·
0

Data Processing Is Becoming a GPU Workload

anyscale.com·2 pts·robertnishihara·
0

67% Cost Savings with PD Disaggregation Using Ray and vLLM on AMD MI325X

anyscale.com·4 pts·robertnishihara·
0

Major upgrades to Ray Serve: 88% lower latency and 11.1x higher throughput

anyscale.com·2 pts·robertnishihara·
1

SkyRL brings Tinker to your GPUs (2025)

novasky-ai.notion.site·24 pts·robertnishihara·
5

vLLM large scale serving: DeepSeek 2.2k tok/s/h200 with wide-ep

blog.vllm.ai·147 pts·robertnishihara·
54

Massively Parallel Agentic Simulations with Ray

anyscale.com·2 pts·robertnishihara·
0

Deploy DeepSeek‑R1 with VLLM and Ray Serve on Kubernetes

anyscale.com·1 pts·robertnishihara·
0

An Open Source Stack for AI Compute: Kubernetes and Ray and PyTorch and VLLM

anyscale.com·1 pts·robertnishihara·
0

Native LLM APIs in Ray Data and Ray Serve

anyscale.com·2 pts·robertnishihara·
0

Joins and Hash-Shuffle in Ray Data

anyscale.com·3 pts·robertnishihara·
0

AsyncFlow: An Asynchronous Streaming RL Framework for LLM Post-Training

arxiv.org·4 pts·robertnishihara·
0

Open Source RL Libraries for LLMs

anyscale.com·1 pts·robertnishihara·
0

Large-Scale Deployment of Ray in Tencent's Weixin AI Infrastructure

anyscale.com·2 pts·robertnishihara·
0

Uv and Ray: Pain-Free Python Dependencies in Clusters

anyscale.com·44 pts·robertnishihara·
10

Roll: Reinforcement Learning Optimization for Large-Scale Learning

github.com·1 pts·robertnishihara·
0

An Open Source Stack for AI Compute: Kubernetes and Ray and PyTorch and VLLM

anyscale.com·1 pts·robertnishihara·
0

Uv and Ray: Pain-Free Python Dependencies in Clusters

anyscale.com·1 pts·robertnishihara·
0

Ray Batch Inference at Pinterest (Part 3)

medium.com·1 pts·robertnishihara·
0

Direct Preference Optimization with Synthetic Data on Anyscale

anyscale.com·1 pts·robertnishihara·
0

Building an LLM Router for High-Quality and Cost-Effective Responses

anyscale.com·1 pts·robertnishihara·
0

Ray Infrastructure at Pinterest

medium.com·1 pts·robertnishihara·
0

Lessons from training a Stable Diffusion model on 2B images

anyscale.com·5 pts·robertnishihara·
0

Canva Built a Modern AI Platform Using Anyscale

anyscale.com·2 pts·robertnishihara·
0

Building RAG-Based LLM Applications for Production

anyscale.com·2 pts·robertnishihara·
0

Fine-tuning LLMs for longer context and better RAG systems

anyscale.com·1 pts·robertnishihara·
0

Two-day hands-on RAG Bootcamp for developers

twitter.com·2 pts·robertnishihara·
0

RAG at Scale: 10x Cheaper Embedding Computations with Anyscale and Pinecone

anyscale.com·1 pts·robertnishihara·
0

Comparing LLM Performance: Introducing the Open Source Leaderboard for LLM APIs

anyscale.com·2 pts·robertnishihara·
0

LLMPerf Leaderboard

github.com·5 pts·robertnishihara·
0
Page 1 of 3Next →