HNHacker News
TopNewBestAskShowJobs

deet

768 karma · joined January 20, 2012

ex-Apple AI/ML. Second-time founder working on Kinelo

If you think in long timeframes, want to build for the future of software, join us at kinelo.com

Generally happy to chat (except recruiters...): keith [at] kinelo.com

submissionscomments
deet··on Introducing The Model Context Protocol
Landing page link is in my bio

We've been keeping quiet, but I'd be happy to chat more if you want to email me (also in bio)

deet··on Model Context Protocol
My team and I have a desktop product with a very similar architecture (a central app+UI with a constellation of local servers providing functions and data to models for local+remote context)

If this protocol gets adoption we'll probably add compatibility.

Which would bring MCP to local models like LLama 3 as well as other cloud providers competitors like OpenAI, etc

deet··on Ask HN: Who is hiring? (July 2024)
Avy (https://www.avy.ai) | Multiple Roles | Salt Lake City, UT | REMOTE (USA) or ONSITE

We are an early-stage, well-funded, stealth startup making humans and computers work together more efficiently. Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- Integrations and server engineer (Go/TypeScript/Python connections to data sources, and data syncing) with some devops responsibilities

- Generalist AI/ML engineer (writing agent code, RAG, prompt engineering, etc)

We're distributed but expect travel for regularly scheduled on-site, in-person work in SLC, with future presence in New York City.

Please visit https://avy.breezy.hr

deet··on Ask HN: Who is hiring? (June 2024)
Avy (https://www.avy.ai) | Multiple Roles | Salt Lake City, UT | REMOTE (USA) or ONSITE

We are an early-stage, well-funded, stealth startup making humans and computers work together more efficiently. Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- Senior Applied AI/ML engineer (including LLM fine-tuning, search/retrieval systems, and various vision and NLP tasks)

- Generalist AI/ML engineer (writing agent code, RAG, prompt engineering, etc)

- Integrations and server engineer (TypeScript/Go/C++ connections to data sources, and data syncing)

We're distributed but expect travel for regularly scheduled on-site, in-person work in SLC, with future presence in New York City.

Email jobs@avy.ai or visit https://avy.breezy.hr (not all positions posted there yet)

deet··on Ask HN: Who is hiring? (April 2024)
Avy (https://www.avy.ai) | Multiple Roles | Salt Lake City, UT | REMOTE (USA) or ONSITE

We are an early-stage, well-funded, stealth startup making humans and computers work together more efficiently. Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- Senior Applied AI/ML engineer (including LLM fine-tuning, search/retrieval systems, and various vision and NLP tasks)

- Marketing (in the "growth hacker" spirit) -- if your dream is to launch the fastest-growing B2B SaaS product ever, we want to talk with you.

We're distributed but expect travel for regularly scheduled on-site, in-person work in SLC, with future presence in New York City.

Email jobs@avy.ai or visit https://avy.breezy.hr (not all positions posted there yet)

deet··on Ask HN: Who is hiring? (March 2024)
Avy (https://www.avy.ai) | Multiple Roles | Salt Lake City, UT | REMOTE (USA) or ONSITE

We are an early-stage, well-funded, stealth startup making humans and computers work together more efficiently. Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- Senior Applied AI/ML engineer (including LLM fine-tuning, search/retrieval systems, and various vision and NLP tasks)

- Generalist AI/ML engineer (writing agent code, RAG, prompt engineering, etc)

- Marketing (in the "growth hacker" spirit) -- if your dream is to launch the fastest-growing B2B SaaS product ever, we want to talk with you.

- MacOS (Swift, Objective-C, C/C++, etc) at the Senior and Staff levels. iOS experience is OK.

We're distributed but expect travel for regularly scheduled on-site, in-person work in SLC, with future presence in New York City.

Email jobs@avy.ai or visit https://avy.breezy.hr (not all positions posted there yet)

deet··on Ask HN: Who is hiring? (February 2024)
Avy (https://www.avy.ai) | Multiple Roles | Salt Lake City, UT | REMOTE (USA) or ONSITE

We are an early-stage, well-funded, stealth startup making humans and computers work together more efficiently. Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- Generalist AI/ML engineer (writing agent code, RAG, prompt engineering, etc)

- Senior Applied AI/ML engineer (including LLM fine-tuning, search/retrieval systems, and various vision and NLP tasks)

- Marketing (in the "growth hacker" spirit) -- if your dream is to launch the fastest-growing B2B SaaS product ever, we want to talk with you.

- MacOS (Swift, Objective-C, C/C++, etc) at the Senior and Staff levels. iOS experience is OK.

We're distributed but expect travel for regularly scheduled on-site, in-person work in SLC, with future presence in New York City.

Email jobs@avy.ai or visit https://avy.breezy.hr (not all positions posted there yet)

deet··on Ask HN: Who is hiring? (January 2024)
Weird! Thanks for letting us know. It's just sending to MailChimp but we'll take a look
deet··on Ask HN: Who is hiring? (January 2024)
Avy (https://www.avy.ai) | Multiple Roles | Salt Lake City, UT | REMOTE (USA) or ONSITE

We are an early-stage, well-funded, stealth startup making humans and computers work together more efficiently. Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- Generalist AI/ML engineer (writing agent code, RAG, prompt engineering, etc)

- Senior Applied AI/ML engineer (including LLM fine-tuning, search/retrieval systems, and various vision and NLP tasks)

- Marketing (in the "growth hacker" spirit) -- if your dream is to launch the fastest-growing B2B SaaS product ever, we want to talk with you.

- MacOS (Swift, Objective-C, C/C++, etc) at the Senior and Staff levels. iOS experience is OK.

We're distributed but expect travel for regularly scheduled on-site, in-person work in SLC, with future presence in New York City.

Email jobs@avy.ai or visit https://avy.breezy.hr (not all positions posted there yet)

deet··on Ask HN: Who is hiring? (December 2023)
Avy (https://www.avy.ai) | Multiple Roles | Salt Lake City, UT | REMOTE (USA) or ONSITE

We are an early-stage, well-funded, stealth startup making humans and computers work together more efficiently. Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- MacOS (Swift, Objective-C, C/C++, etc) at the Senior and Staff levels

- Applied ML (including LLM fine-tuning and various vision and NLP tasks)

- Marketing (in the "growth hacker" spirit)

We're distributed but expect travel for regularly scheduled on-site, in-person work in SLC, with future presence in New York City.

Email jobs@avy.ai or visit https://avy.breezy.hr

deet··on Ask HN: Who is hiring? (November 2023)
Avy | Multiple Roles | Salt Lake City, UT | REMOTE (USA) | ONSITE

We are an early-stage, stealth startup making humans and computers work together more efficiently.

Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- Design (UI/UX)

- MacOS (Swift, Objective-C, C/C++, etc)

- Applied ML (including LLM fine-tuning, various NLP tasks)

- Marketing (in the "growth hacker" spirit)

We're distributed but expect travel for regular on-site, in-person work in SLC, with future presence in New York City.

jobs@avy.ai

deet··on Ask HN: Who is hiring? (October 2023)
Avy | Multiple Roles | Salt Lake City, UT | REMOTE (USA) | ONSITE

We are an early-stage, stealth startup making humans and computers work together more efficiently.

Experienced team from Apple AIML, Bose, Amazon, and other great companies.

We're hiring for:

- Design (UI/UX)

- MacOS (Swift, Objective-C, C/C++, etc)

- Applied ML (including LLM fine-tuning, various NLP tasks)

- Part-time/contract roles for TypeScript and Python, and server work (Go)

- Marketing (in the "growth hacker" spirit)

We're distributed but expect travel for regular on-site, in-person work in SLC, with future presence in New York City.

jobs@avy.ai

deet··on Unexpected benefits of sun exposure on skin
It's surprising this article doesn't significantly mention sunlight's non-UV components, like infrared light. There appears to be increasing evidence to support its ability to influence various processes within the body to positive effect.

MedCram has some easily digestible reviews on the topic:

- The ability for light to influence glucose metabolism: https://www.youtube.com/watch?v=6Win49aeh8A

- Light's effect on immune response and other processes: https://www.youtube.com/watch?v=5YV_iKnzDRg

deet··on Fine-tune your own Llama 2 to replace GPT-3.5/4
Azure GPT 4 is already available in: Australia East, Canada East, East US, East US 2, France Central, Japan East, Sweden Central, Switzerland North, UK South (https://learn.microsoft.com/en-us/azure/ai-services/openai/c...)
deet··on Show HN: LlamaGPT – Self-hosted, offline, private AI chatbot, powered by Llama 2
It's possible to extend the effective context window of many OSS models using various techniques. The Llama-related models and others there's a technique called "RoPE scaling" which allows you to run inference over a longer context window than the model was originally trained for. (This reddit post help highlight this fact: https://www.reddit.com/r/LocalLLaMA/comments/14lz7j5/ntkawar...)

But even at 100K, you do eventually run out of context. You would with 1M tokens too. 100K tokens is the new 64K of RAM, you're going to end up wanting more.

So techniques like RAG that others have mentioned are necessary in the end at some point, at least with models that look like they do today.

deet··on Show HN: LlamaGPT – Self-hosted, offline, private AI chatbot, powered by Llama 2
The computing power is definitely not out of reach of mere mortals. I'm working on software that does this for emails and common documents, generating a hybrid semantic (vector) and keyword search system over all your data, locally.

The computing power we're requiring is simply what's available in any M1/M2 Mac, and the resource usage for the indexing and search is negligible. This isn't even a hard requirement, any modern PC could index all your emails and do the local hybrid search part.

Running the local LM is what requires more resources, but as this project shows it's absolutely possible.

Of course getting it to work *well* for certain use cases is still hard. Simply searching for close sections of papers and injecting them into the prompt as others have mentioned doesn't always provide enough context for the LM to give a good answer. Local LMs aren't great at reasoning over large amounts of data yet, but getting better every week so it's just a matter of time.

(If you're curious my email is in my profile)

deet··on The elite's war on remote work has nothing to do with productivity
Can't edit the post now, but if the second link is broken it was meant to point to this: https://www.bloomberg.com/news/articles/2023-03-30/wall-stre...

Here's another related one: https://www.forbes.com/sites/jackkelly/2022/02/17/new-york-c...

deet··on The elite's war on remote work has nothing to do with productivity
The article is not well argued, yes.

But there is direct evidence that politicians are influencing corporations to alter return to office policies:

- Mayor of SF asking businesses to pledge to implement RTO policies: https://sfist.com/2022/03/03/mayor-breed-would-like-you-back...

- Mayor of NYC basically doing the same: https://archive.is/si6xd

deet··on Show HN: liteLLM Proxy Server: 50+ LLM Models, Error Handling, Caching
Is user auth (and tracking token spend) within scope of this or is that better handled at a layer in front of this?
deet··on Zero Motorcycles makes its service manuals free
I pay like $8/month for access to my 2021 Triumph’s service manual.

As far as I know they keep it as an interactive website so that PDFs are hard to make.

I haven’t found a good copy yet but it’s also possible that I’m just not as good at searching for pirated content as back in the old torrenting days (or perhaps just less motivated)

deet··on GPT-4 is getting worse over time, not better
In looking at the paper's continuous mention of "ChatGPT" and the repo README's statement that "You don't need API keys to get started" .. are we sure they weren't using type of tools to talk to the ChatGPT API (via a session token, etc) vs the OpenAI API? I do agree they talk about the API in the paper a lot but I don't see an exact methods statement that they directly accessed the non-ChatGPT API anywhere, unless I'm missing it,
deet··on Launch HN: Credal.ai (YC W23) – Data Safety for Enterprise AI
Looks awesome and will make many enterprises feel more comfortable using AI.

I suspect your intuition about moving emphasis from redaction to unified access control and audit logging over time is right.

The "AI Chief of Staff" sounds interesting though -- can you share a bit more about what you showed to companies and received lukewarm response to?

deet··on GGML – AI at the Edge
The parent is saying that "fine tuning", which has a specific meaning related to actually retraining the model itself (or layers at its surface) on a specialized set of data, is not what the GP is actually looking for.

An alternative method is to index content in a database and then insert contextual hints into the LLM's prompt that give it extra information and detail with which to respond with an answer on-the-fly.

That database can use semantic similarity (ie via a vector database), keyword search, or other ranking methods to decide what context to inject into the prompt.

PrivateGPT is doing this method, reading files, extracting their content, splitting the documents into small-enough-to-fit-into-prompt bits, and then indexing into a database. Then, at query time, it inserts context into the LLM prompt

The repo uses LangChain as boilerplate but it's pretty easily to do manually or with other frameworks.

(PS if anyone wants this type of local LLM + document Q/A and agents, it's something I'm working on as supported product integrated into macOS, and using ggml; see profile)

deet··on Ask HN: What's the best self hosted/local alternative to GPT-4?
In our experimentation, we've found that it really depends what you're looking for. That is you really need to break down down evaluation by task. Local models don't have the power yet to just "do it all well" like GPT4.

There are open source models that are fine tuned for different tasks, and if you're able to pick a specific model for a specific use case you'll get better results.

---

For example, for chat there are models like `mpt-7b-chat` or `GPT4All-13B-snoozy` or `vicuna` that do okay for chat, but are not great at reasoning or code.

Other models are designed for just direct instruction following, but are worse at chat `mpt-7b-instruct`

Meanwhile, there are models designed for code completion like from replit and HuggingFace (`starcoder`) that do decently for programming but not other tasks.

---

For UI the easiest way to get a feel for quality of each of the models (or, chat models at least) is probably https://gpt4all.io/.

And as others have mentioned, for providing an API that's compatible with OpenAI, https://github.com/go-skynet/LocalAI seems to be the frontrunner at the moment.

---

For the project I'm working on (in bio) we're currently struggling with this problem too since we want a nice UI, good performance, and the ability for people to keep their data local.

So at least for the moment, there's no single drop-in replacement for all tasks. But things are changing every week and every day, and I believe that open-source and local can be competitive in the end.

deet··on Hard stuff when building products with LLMs
I suspect you're right for how people are using and deploying LLMs now: hacking all kinds of functionality out of a text-completion model that, although it encodes a ton of data and some reasoning, is fundamentally still a text completion model and when deployed via commercial APIs like today without fine tuning, are not flexible beyond prompt engineering, chaining, etc. make possible.

But I think we've only scratched the surface as to what LLMs fine-tuned on specific tasks, especially for abstract reasoning over narrow domains, could do.

These applications possibly won't look anything like the chat interfaces that people are getting excited about now, and fine-tuning is not as accessible as prompt engineering. But there's a whole lot more to explore.

deet··on SFPD obtained live access to business camera network in anticipation of protest
A convincing case against this probably can't be made.

But that's the problem. Each incremental step towards more surveillance, less privacy, and more potential for government abuse is perfectly justifiable and seems reasonable.

But then temporary turns to permanent. And the "significant events" restriction gets dropped. And 450 cameras from local businesses turn into thousands from others, or Rings, or Teslas, or whatever.

And then manual monitoring by humans turns into AI-powered monitoring. And the looking at cameras gets combined with location data.

The point is, each step is reasonable. But who knows where it goes? We have no idea, nor do we have any idea who will be on the other side watching or what their agenda will be in 5 years, 20 years, or 100 years.

So it's important to stay vigilant of any incremental privacy incursion or expansion of government power. It doesn't mean saying No necessarily, but being aware and cautious.

deet··on Bringing the power of AI to Windows 11
There are two things to note here:

1. A cross-application connectivity layer that pipes data and actions between apps

2. A natural language interface to control #1

Thinking about them separately is useful, because although chat is the new UI hotness, #1 is valuable on its own and the two can potentially be deployed separately.

As presented here, I suspect the natural language interface will be faster and easier than buttons for operating the cross-app layer for complex queries, but potentially slower than operating buttons for simple things (like "start dark mode").

But personally, I believe #1 combined with some AI context awareness is more powerful of the features.

...

And btw, I left Apple last year to build a local-first and developer-extensible assistants for the Mac that's pretty similar. If this interests you, would love to chat (email in profile, as well as a waitlist).

deet··on Following UK antitrust order, Meta sells Giphy to Shutterstock for $53M
This is the second divestiture of an acquisition in two weeks by Meta.

Last week they spun out Kustomer https://www.kustomer.com/blog/new-chapter-standalone-company...

deet··on It’s time to embrace slow productivity (2022)
Comments here so far seem to be missing Newport's key point (in the second half of the article)...

For many modern (knowledge work) jobs, it's volume of tasks, not duration, that seems to induce burnout.

For instance, if you have 1 central task for the week--say, write a report--but there are ten subtasks (hold 5 meetings to prepare, read 3 background papers, ...), and then each of those have a bunch of subtasks (Slack each 5 meeting attendees 10 times to coordinate and schedule, get interrupted when reading each paper so have to resume X times each) ... this 1 central task can easily be dozens or hundreds of subtasks.

And then each coworker is doing the same thing, adding tasks to your load (you have to read their messages, respond to their emails, etc) in a multiplicative way. Newport calls it the "Hive Mind" in one of his books. The number of total tasks, from very small to very large, each individual has to accomplish in a week ends up far far greater than expected on the surface. And adding to that, they come in in an unpredictable way.

All this adds up to burnout, not just the number of hours. It's the intensity, and the unpredictability.

I've experienced this myself. I was at one of the FAANGs, constantly bombarded with new tasks, and felt burnt out. Now, I'm at a startup--very much inspired by Cal Newport, and using AI and context awareness to make teams operate together more effectively (see bio)--and I'm working far more hours per week than before, but with less interruptions and less distraction, I'm able to focus and feel far less burnout despite the increased hours.

All this to say, we really need to rethink how we work, not just how much.

deet··on Apple Restricts Employee Use of ChatGPT, Joining Other Companies Wary of Leaks
Apple's security policy already prohibited putting confidential or proprietary information into any not-explicitly-approved, externally-hosted service. So I'm guessing this wasn't a "ban" so much as reminding people of this policy and that ChatGPT is still unapproved.

The reasons go behind using data as training. Submissions to servers end up in logs, databases, temp files... who knows. And a company like Apple wants to not only ensure that the data is explicitly used by the receiving party, but also not inadvertently made accessible to others via poor security procedures, since their data is such a high value target.

← PreviousPage 2 of 6Next →