Hacker Newsnew | past | comments | ask | show | jobs | submit | manapause's commentslogin

Model + Harness = Agent

Shame us what you lose when your turn to someone or something you trust for empathy and instead receive the opposite. Our body-personality would not produce venom if we didn’t need it to survive.

AMEN. If AI has made a developer into a curmodeon, I honestly would love to meet them.

The fact that in the 1950s and 60s our best and brightest from around the world dreamt of working for jonas salk or joining the space program.

for the last 30 years that same level of talent was aggressively pursued to give us Google Adwords, complex insurance-backed financial products, a corporate culture of feckless cowards in leadership, the best of us never stop interviewing. literally an epidemic of dystopian psycho-afflications due to doomscrolling that threatens to devolve future generations thanks to weaponized dopamine manipulation, advertising so pervasive that it could bring down entire goverments (or companies, which are now bigger), or an entire adult generation of intelligent loners who may never own a house...

But AI, AI is what people are mad at?

After 25 years in systems engineering, I was in a state of ennui with technology. I had loved building computers, hacking systems, gaming with my friends so much as a child I built a career out of it.

I worked at a large technology firm where we had a product release bottleneck due to 1 UX engagement artist who's talents were so in-demand we had multiple products backed up because this person had to give birth to a journey that wow, REALLY engaged you

IF a site that sold books, Adwords, Facebook, Twitter toxic-celebrity culture, an infatuation with ourselves and our devices...what if the emperor really has no clothes? what if we are all really lame, lonely bullies to whom the accumulation of wealth comes a cost of empathic and emotional intelligence and a desire to be non-human...

IS THAT THE GOAL? What are we doing here? Creating Optimus Prime, Agent Smith, or Hedonism bot? Maybe all 3 of them fighting for the future? Sold, I'll take it. I fight for Hedonism bot.

nothing AI could do could make me more upset that what humanity has done to itself over the past 30 years. If you are a subject matter expert in technology, NOW is your time.

I just read that the use of dating apps has decreased considerably.

Scammers taking advantage of the weak is more difficult when Joe-luddite asks an advertisement-free AI and eschews wading through a pagerank advertising platform or asking their home amazon-wiretap scams to find the truth themselves - a truth that they can engage with that is somewhat advertising free (for now?)

If you are angry about AI, then go ahead and leave.


If you have both ollama and llama.cpp installed on your MBP, and are interested in utilizing your local model as a cost-saving preprocessor for your frontier models, consider using https://github.com/Standard-Pentest/kultivait!

"kultivait init --setup" to evaluate your machine's hardware, access to frontier cli-tools, download models, and create a proxy for ollama to get started. works great with opencode!


Note: I am replying to my own post to say that Kultivait llama.cpp functionality is currently broken but under active development.

Ollama proxy and model selection does work. If you run into issues or have any questions please feel free to reach out.


I also use it for a chatbot assistant for customer onboarding. Gemini deserves some kudos for their willingness to allow entry level subscriptions access to API-keys to build solutions with.


This is so true. Every junior sysadmin I have trained over the years (including myself) has had a “are we being attacked?!” moment when tasked with WAF report analysis, monitoring fail2ban logs, etc.

Monitoring WAN traffic really gets the paranoia juices flowing.


I remember when you could stand up a website and no bots would scrape it or scan it. It was a lovely time. No one had firewalls or antivirus and things were working fine until the worms and viruses started coming. You could be confident that your guests were real, so much so we had guest counters on many public sites.


You still can.

Just build your website yourself as deep in the stack as you can instead of piling up 50 abstractions on top of each other. Some decisions like having your page be accessible by IP can only happen if you use technology like generic http servers (like apache or nginx) from the 2000s instead of implementing the lower stacks and actually thinking about whether that makes sense for a second.

If when you build a website or a backend, your server responds to requests by IP address (for example), you are building a bottom 90% product, and considering most software markets are super top-heavy, (say 1% win), that's ngmi land.


Remember when you had to submit a request for google to scan your site?


What about at the device level?

“You must be this tall to ride this ride”

“ you must be 18 to own an iPhone 18+ “

I apologize for the drive-by question, and I appreciate your takes!


This would run into the same deal as VINs in the real world being tied to licenses, or serial numbers to guns. But the car equivalent of VMs/open hardware/custom firmware (imagine a $7 pi zero flashed with lineageOS “overage phone”) then becomes equivalent to a gun without a serial number, and suddenly open source/hardware people are felons and there is insane amount of control on hardware and software like it’s 1982.

This assumes that the government would be able to verify independently a phone serial number so that people’s IDs aren’t leaked. If not, then you’re back to the same thing as before since “drivers licenses” are stored by sites and shared around with advertisers


Can confirm, my experience in “loop engineering” was “this is neat” for 45 minutes until a daily ration of tokens was evaporated. The quadratic cost trap is prohibitive to experimentation.

As a localLLM evangelist, I am hopeful this will bring more attention to the joys of rolling your own sovereign AI.


Yeah, i'm hoping that gets smoother. I've been experimenting with omlx and opencode on my m5x64gb and keep running into issues w/ Qwen3.6-35B-A3B-MLX-8bit exceeding it's memory limit at the most inopportune times. Playing with 12B gemma4 (8bit) more today.

Maybe I should be aiming for something targeting 48gb of memory?


It depends what your goals are and what you are using it for. This space is fluid and my answer last week would be different than my answer today! That said there’s no substitute for hard work, here are some resource to get you up to up to speed:

https://carteakey.dev/blog/local-inference/local-llm-optimiz...

https://botmonster.com/ai/self-hosted-ai-agent-frameworks-20...

Personally I find myself swapping models depending if I am engaged in “trad-development” vs building agentic probes or apps involving imagery. Tailscale the LLM to your deployments and ta-da!


The irony in AI triggering societal collapse due to gross economic malfeasance is just fun to think about.

If AI was around in the early 2000s Countrywide.ai would have been a thing.


The more sensational the headline the less I believe that the authors were present in technology 15-20+ years ago. People forget that Reddit used to be 2 parts programmer-humor 1 part snuff.

Show me an abliterated frontier model that is able to breakthrough the surrounding supporting models and actually hold state to produce contraband and I’ll gladly supply my personal image making making a silly face in a compromising position if it wouldn’t make the testers feel better.

Do they need to be tested like this? Yes. But it would take the carbon footprint of a commuter air terminal and the land rights of am small town in the high Sierras …. all converted settlers of Catan style into tokens …. just to lobotomize a fine tuned model to get close.

That said I appreciate the work you’re doing


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: