HNHacker News
TopNewBestAskShowJobs

amarble

318 karma · joined March 20, 2021

submissionscomments

Evaluating AI Agents as Products

marble.onl·1 pts·amarble·
0

Evaluating AI Agents as Products

marble.onl·2 pts·amarble·
0

Updating IP Regulations for AI Distillation

marble.onl·1 pts·amarble·
0

Updating IP Regulations for AI Distillation

marble.onl·2 pts·amarble·
0

LLMs are still just low code / no code software

marble.onl·5 pts·amarble·
1

LLMs are still just low code / no code software

marble.onl·2 pts·amarble·
0

There is minimal downside to switching to open models

marble.onl·407 pts·amarble·
330

AI doesn't replace white collar work

marble.onl·67 pts·amarble·
104

I paid $170 and all I got was this demo

marble.onl·3 pts·amarble·
1

I paid $170 and all I got was this stupid demo

marble.onl·1 pts·amarble·
0

Task-free intelligence testing of LLMs

marble.onl·69 pts·amarble·
22

Task-free intelligence testing of LLMs

marble.onl·1 pts·amarble·
0

Intelligence is not just about task completion

marble.onl·1 pts·amarble·
0

If You Meet ET in Space, Kill Him (2024)

nautil.us·3 pts·amarble·
0

Intelligence is not just about task completion

marble.onl·3 pts·amarble·
0

Show HN: Gen AI Writing Showdown

writing-showdown.com·2 pts·amarble·
0

Ifrro member Kopinor signs agreement on newspaper content for AI in Norway

ifrro.org·2 pts·amarble·
0

Comparing language model performance on creative writing transformations

writing-showdown.com·2 pts·amarble·
0

Eminembench

marble.onl·1 pts·amarble·
0

Promptware Attacks Against LLM-Powered Assistants in Production

sites.google.com·1 pts·amarble·
0

Managing LLM application performance through code standards

marble.onl·2 pts·amarble·
0

Catching Claude Cheating

marble.onl·1 pts·amarble·
0

Catching Claude Cheating

marble.onl·1 pts·amarble·
1

Scanning AI application code for vulnerabilities and performance issues

marble.onl·3 pts·amarble·
0

Show HN: A static scanner for LLM app code

github.com·6 pts·amarble·
1

Scanning AI application code for vulnerabilities and performance issues

marble.onl·2 pts·amarble·
0

The Model Trust Score: The Framework for Strategic Enterprise AI Model Selection

credo.ai·2 pts·amarble·
0

Evals are not all you need

marble.onl·58 pts·amarble·
12

An AI Cyber Incident in Plain Sight

marble.onl·2 pts·amarble·
0

AI agent using Anthropic's tool calling and the Pandas Python library

github.com·2 pts·amarble·
0
Page 1 of 2Next →