286 karma · joined September 12, 2010
Dollhouse Research — composable AI tools, mostly open source (dollhouseresearch.com):
- DollhouseMCP: Open-source AI customization with AI safety as at it's core. Turn prompts into modular building blocks you can create, activate, combine, version, and share — Personas, Skills, Templates, Agents, Memories, and Ensembles. (dollhousemcp.com)
- Dollhouse Collection: community catalog of those elements (collection.dollhousemcp.com)
- MCP-AQL: a protocol spec that gives MCP tool calls semantic meaning — Create/Read/ Update/Delete/Execute endpoints that cut token overhead and make safe-vs-destructive operations explicit at the protocol layer (mcpaql.com)
- AILIS: an OSI-inspired, layered model that gives AI systems a shared vocabulary, so teams can tell a routing problem from a tool-invocation, memory, or UX one
- Also check out Merview.com, a local-first Markdown and Mermaid editor and viewer that runs entirely in your browser, no login
Elemental Surveys (elementalsurveys.com):
- Fast-turn market research — turn a brief into a QA'd questionnaire, analysis, charts, and a branded report in hours instead of weeks
Some Previous Projects:
- Created Tomorrowish.com a social media DVR used by Hulu, Fox, Turner Broadcasting, and more
- Cofounder and CEO of TVDuffle a Travel and booking app for concert goers and sports fans.
- Cofounder Biostrut 3D Printed organic tissue scaffold
ping me at my firstname@myfullname.com
My personal blog can also be found at http://mickdarling.com
It only has five CRUDE endpoint: Create, Read, Update, Delete, and Execute using a GraphQL-like structure for tool calling of the operations within the endpoints. It's very efficient, and robust. there's all kinds of exemplar tools and components to make adapters for any MCP server. You don't even need to rewrite your own MCP server. Just create an adapter for it.
All open source at MCPAQL.com
And, they can say that for anybody at any time, and you'll never know why, and there's no way to prove it.
Everyone needs a flight data recorder to prove... "here's what I was actually doing and why it was not distillation." And now you're having to prove your innocence instead of them having to prove you're guilty, and really at the end of the day, it's just the model being stupid that they're protecting themselves from.
They obviously put their best model on the job to build that.
----------------------
Fable 5: Our most capable model yet Our newest model tackles your biggest challenges with fewer check-ins needed.
• <b>Included in your plan limits until Jun 22</b><br><br>Fable takes 2× the usage of Opus. • <b>Switch models when a message is flagged</b><br><br>When safety measures flag a message, automatically switch to a different model to keep chatting. When off, your chat will pause instead. <a href="https://support.claude.com/en/articles/15363606" target="_blank" rel="noopener noreferrer">Learn more</a>
There's a lot of room for improving the smaller models at many levels of the stack.
Then I looked at the usage and it said I had used 95% of my Claude design usage for the week!
This isn't a real tool. This is a plaything, if that's what they're providing as examples.
For me, it is having a document and interrogating it. Maybe having many sets of documents about a whole category of information. Getting the bullet points. getting the high level and then interrogating and digging down and being able to get bubbled up information as I need it.
That is the learning style that matches how I learn.
I have never been able to skim, so reading a large document WILL teach me that topic, but getting through that doc is tough.
I can dump a very large set of docs in a reader that lets me interrogate the whole data set and I can fly through looking for what is interesting to me, and what I may need, and along the way I will likely dive into other parts too. Asking questions keeps my hyperfocus active.
I think it is just a different style. I have synesthesia and a hard time not working on three to five things at once. I am use to knowing I learn differently than others.
When it comes to agents' tasks, I tend to focus on things that I couldn't do before without automated agents, at least at the going price.
The kind of automation I'm doing is more like building a set of agents to generate marketing surveys for me. They take free form input from me and my project. They aren't particularly sexy but they go off and do something valuable that I literally would never pay for at the prices that they are normally.
If I had a nice CI/CD workflow that was built into GitHub rather than rolling my own that I have running locally, that might just make it a little more automatic and a little easier.
As for the domain, this is the same account that has been hosting Github projects for more than a decade. Pretty sure it is legit. Org ID is 9,919 from 2008.
We have mechanisms for ensuring output from humans, and those are nothing like ensuring the output from a compiler. We have checks on people, we have whole industries of people whose whole careers are managing people, to manage other people, to manage other people.
with regards to predictability LLMs essentially behave like people in this manner. The same kind of checks that we use for people are needed for them, not the same kind of checks we use for software.
It looks far too risky to use, even if I have it sequestered in its own VM. I'm not comfortable with its present state.
Whether you had anything to do with it or not, I have no idea. And, since you didn't follow best practices and tell me directly rather than trying to score points here, there's really no way of knowing whether you're the one who caused the problem in the first place.
I built a new site without Wordpress. That took in less than a day.
I don't imagine you will alter your behavior to align with general best security practices anytime soon.
There are a lot of really bad human developers out there, too.
The failure modes are just too rough for most people to think about until it's too late.
https://merview.com with full source code at https://github.com/mickdarling/merview
I can tell you that a good number of the design drawings for the higher floors in the Venetian resort in Las Vegas were assembled with AutoLisp scripts. The scripts I created grabbed components from other drawings that were already made to assemble a first pass set of drawings for floors that hadn't been fully designed yet, since the floors all had components of other floors.
They were still in the design process for the upper floors, while the lower floors had already been finished and they were moving up the building.
I 100% agree, and I own very nice LG TVs. They are not connected to the internet. They each have an Apple TV and that is their only way that they get video, and can't send data out.