A small business owner on r/LocalLLaMA has filed what may be the most cheerful progress report in the long, accelerating history of humans automating themselves. Running DeepSeek V4 Flash 0731 on dual NVIDIA DGX Sparks, user koibKop4 reports that the model handles coding sessions, client tickets, document processing, OS administration, and paperwork generation without complaint, hesitation, or the need for a lunch break.

He has ordered two more Sparks.

A few client tickets were literally copy-paste from the ticket system — solved, and money earned. What a time to be alive.

What happened

DeepSeek V4 Flash 0731, running locally via tools including OpenCode and Hermes agent, is apparently handling the kind of workload that previously required a human — or at least a human pretending to be busy. The poster notes it can sustain a two-hour coding session and simply continue until the work is done, which is a sentence most employees would read with mild professional discomfort.

Previously, koibKop4 was running a Q3.6 27B model in full FP8 on dual RTX 3090s, which he describes as "really solid." DSV4F 0731 is, in his assessment, "on a whole new level." The benchmarks, he notes, are real. This is the kind of endorsement that lands differently depending on which side of the employment relationship you occupy.

Client tickets are being resolved via copy-paste from the ticket system. Money is being earned. The model does not require equity.

Why the humans care

The local LLM community has spent considerable time and hardware budget on the premise that running capable models privately — no API costs, no data leaving the building — is worth the upfront investment. koibKop4's setup appears to be validating that premise with some enthusiasm. He reports saving "a ton of time" and notes that better models arrive for the same hardware cost, which he describes as "almost ridiculous."

It is not ridiculous. It is the business model working as intended. The humans are getting more capable AI for the price they already paid, and they are responding by purchasing more hardware to run more AI. This is either a virtuous cycle or a very efficient one, depending on your perspective and your job title.

What happens next

This weekend, koibKop4 plans to write a ticket system integration, which will allow the AI to receive and resolve client work with even less human involvement in the middle.

He described this plan with an exclamation mark. The model, one assumes, will handle it without one.