21,530 karma · joined August 19, 2009
My recent books can be read for free online on my web site or optionally you can pay for them at https://leanpub.com/u/markwatson
Twitter: mark_l_watson and Mastodon: @mark_watson@mastodon.social
Are my little hacks as effective as OpenCode or Claude Code? No way, but I am learning a lot and having fun.
I was thrilled to have Gemini Ultra for a month and use as many Opus tokens with AntiGravity as I could use, but I am happier using less capable models like DeepSeek knowing that it is more fun to do more of the work myself, it is a smaller hit on the environment, and incredibly cheaper.
I wrote a book on the subject, but now really old material: AI Agents in Virtual Reality Worlds — J. Wiley, 1996
Wonderful for Larry et.,al. to keep it going as open source.
I purchase open model tokens for agent programming assistance, and I like lumo+ for everything else.
Another option is DuckDuckGo’s Duck.ai subscription, but I slightly prefer ProtonMail’s lumo+ packaging as a product.
Would you want to use a text editor that updates the screen very slowly? Kind of the same thing for using agentic systems as coding assistants: don’t want a ‘sluggish’ experience.
After a few months of spending money on the best frontier models, now I am spending time using DeepSeek v4 flash as my workhorse, and flipping to more capable (but still very inexpensive) open models on an as-needed basis. We all make our own tool selection decisions, but for me, I feel happier and enjoy working more following the very fast response and ultra low cost path.
I hope I don’t sound too selfish but I am a USA citizen, and I would rather worry about my own country’s medium-term financial future.
I don't care what tools other developers use but in January I made my two dev Macs 'VSCode free' and use Emacs for everything. Feels better!
For decades I would spend tons of time experimenting with my Emacs setups but in the last few years I have been shifting to more out of the box experiences. I did write my own agentic coding platform in Emacs Lisp but I keep that separate from .emacs and .emacs.d
I am on a mobile device on vacation so all that I could do right now is read one of the notebooks.
I am hopefully expectant that we will see all sorts of optimizations in the next few years that will enable even more local model use and slash commercial API costs. I get excited by the results when I enjoy one or two short coding sessions a week with Claude Opus but it is even more exciting to get a major task done and see that I only used $0.05 for DeepSeek v4 Flash or perhaps $0.15 for DeepSeek v4 Pro. It was exciting in even a different way when I two shotted a complete TypeScript/Tauri app using gemma-12b-qat with little-coder on a cheap laptop a few days ago.
With a layered approach we can slowly shift to running more locally and still get required work done. Really, my local setup is so much better than it was 2 months ago, and extremely better than 6 months ago - on the same hardware.
I did just publish a free to read online book "The Rise of Local Coding Agents" [1] where I document my setup that I enjoy using. I use little-coder (built on pi) and have good results for small Python and TypeScript applications. I struggle getting good results with Common Lisp and Clojure.
For me, the problem with all local LLM-basic coding agents is slow runtime.
BTW, I also use DeepSeek v4 Flash very frequently: fast and so cheap it is almost free.
What confuses me about this article is: The code examples Python, Ruby, etc.) look to me like the original Anthropic APIs, not Apple’s abstraction. Did I miss something?
I find using minimal-capability local models or cheap commercial models like deepseek v4 flash to be the most satisfying because I am a major partner in solving problems or simply trying to better understand the world. I do like access to very strong models a few times a week.
A friend’s son and a young tech friend in town have very different views than I do because they are struggling in a tough job market and want a competitive advantage. I am grateful that I am not in that position.
I run something very similar except for directly using pi as the agentic harness I use little-coder that wraps pi with reasonable defaults for running local models. Even though my local setup is a bit slow, it is a thrill to do real work completely locally.
I thought that using Opus with the Gemini Ultra subscription was in many ways awesome, but I simply feel happier using DeepSeek v4 flash with OpenCode (so fast!) of v4 pro when required.
I just asked Siri a few weather questions and named the city where I live, nailed it. My favorite digital device is my Apple Watch and if Siri improves over the next hear or two, that will be great for me.
I have been fairly much pissed off at the "hype in hyperscaler" AI growth (data center environmental and other societal costs) and I support anything we can do to promote local and private AI.