Tech News
All News AI & ML Architecture DevOps Open Source Programming Team Management Testing & QA Web

Latest News

⚑ Report a Problem

Tech news from the best sources

All topics AI Gear News Tech agents ai api architecture automation beginners career database devchallenge devops javascript llm machinelearning mcp opensource performance productivity programming python react security showdev testing tutorial typescript webdev
All EN RU
EN

Flash Onyx 2.2: teaching a local model law and game feel

(I know, ANOTHER Onyx post) Flash Onyx is the model line behind FLASH , the local-first agent shell I work on. Onyx 2 is gemma4 with a system prompt…

aiollamalocalllmopensource
Dev.to Aug 28, 2026, 16:00 UTC
EN

Moving Scheduled LLM Curation from Cloud APIs to Local Models

Scheduled LLM curation is the least glamorous agent workload you run. A cron job wakes up at 3am, reads a pile of memory, asks a model to dedupe it,…

aiagentslocalllmollamakubernetes
Dev.to Aug 14, 2026, 00:15 UTC
EN

Nine ways to talk to a local model

A nine-tool survey of local-model interfaces on one GPU: what worked, what silently failed, and why what sits between you and the model matters more…

localllmollamallamacppbuildinpublic
Dev.to Aug 11, 2026, 15:56 UTC
EN

Moving Half of Our AI Development to Local LLMs — by Splitting Work by Role, Not by Picking the Biggest Model

Uehara, EarthLink Network Co., Ltd. I build and run more than 20 products by myself, with Claude Code at the core of development. This is a field no…

aillmmachinelearninglocalllm
Dev.to Aug 10, 2026, 23:03 UTC
EN

Teaching a Local AI Agent to Search the Web (Without Lying to You)

Back in early May, someone on the team said something like "let's just add web search, shouldn't take more than a day." Three months and roughly twe…

applesiliconlocalllmllmagentsswift
Dev.to Aug 1, 2026, 14:18 UTC
EN

Why My Local Coding Agent Could Act but Couldn't Finish

There was a month where I blew through my token budget without noticing. Claude Code and Codex, running most of the day, on a codebase I was explori…

llmlocalllmqwenagents
Dev.to Jul 29, 2026, 07:53 UTC
EN

local-llm: A Field Report on Running SOTA Models on Your Own Hardware

The most useful thing in jamesob/local-llm is not the GPU shopping list. It is the fifteen or so BIOS settings, kernel flags, and PCIe hacks that st…

localllmgpuselfhostinginference
Dev.to Jul 20, 2026, 15:02 UTC
EN

LLM Quantization Levels Compared: Q4_K_M vs Q8_0 vs FP16 [2026]

Originally published at kunalganglani.com — read it there for inline code, hero image, and live links. LLM Quantization Levels Compared: Q4_K_M vs Q…

localllmquantizationggufollama
Dev.to Jul 6, 2026, 01:00 UTC
EN

Local AI Agent Browser Extension: Hermes in 120ms

This article was originally published on BuildZn . Everyone's talking about connecting AI to the web, but nobody tells you how to do it privately, w…

aiagentsbrowserextensionlocalllmhermesagent
Dev.to Jun 24, 2026, 07:37 UTC
EN

Cool AI Projects That Failed: The File Integrity Gap

We ship tools that verify software artifacts. We deal with hashes, checksums, and provenance every day. But looking at the local AI landscape, there…

aisecuritylocalllmsoftwarebommodelartifacts
Dev.to Jun 20, 2026, 10:14 UTC
EN

How to Tune llama.cpp --n-gpu-layers: A Practical VRAM Guide (2026)

You already know what --n-gpu-layers does. It moves transformer layers onto your GPU. This post is the next step: how to actually pick the number. I…

localllmllamacppgpuvram
Dev.to Jun 9, 2026, 14:45 UTC
EN

Fitting WhisperX large-v3 + a 24B LLM on one 3090: a reproducible context-capping recipe

This is the technical, reproducible version of a fix I shipped on my own homelab. If you want the narrative version, that's on Medium. This one is t…

homelabollamalocalllmdevops
Dev.to Jun 3, 2026, 03:35 UTC
EN

Run Cursor with a Local Model: Privacy-First AI Coding Without a Subscription

This article was originally published on runaihome.com If you write code for a living, your IDE is now an AI agent — Cursor, GitHub Copilot , Claude…

aicodingcursorlocalllmprivacy
Dev.to Jun 2, 2026, 14:42 UTC
EN

We pre-registered, ran, and verified the macro ablation: information per joule, measured

Maker disclosure: I build Macrokit (Apache-2.0, fully open). This is the data, not a pitch — links and the raw runs at the end. The multi-model benc…

llmlocalllmopensourceai
Dev.to Jun 2, 2026, 09:20 UTC
EN

[Day 9] A local Japanese sentiment AI (BERT) read 8 years of a LINE chat, and the ups and downs surfaced from numbers alone

Intro Day 9. Today is less about model internals and more of a personal experiment: have a local AI analyze the entire chat history with one LINE fr…

localllmaidgxsparkprivacy
Dev.to May 29, 2026, 22:39 UTC
EN

[Day 7] Does Giving an AI More 'Thinking Time' Really Make It Smarter? Training an OpenMythos-Style Mini Model on DGX

[Day 7] Does Giving an AI More "Thinking Time" Really Make It Smarter? Training an OpenMythos-Style Mini Model on DGX Intro Day 7! Reddit kept surfa…

localllmaidgxsparktransformers
Dev.to May 19, 2026, 03:17 UTC
EN

OpenClaw: 13 Errors, $1.50/Month, and an AI Team That Doesn’t Need the Cloud

I run a team of AI agents on a Mac I bought in 2022. They handle my Slack, run research, draft content, monitor infrastructure, and spawn sub-agents…

applesiliconlmstudiolocalllmopenclaw
Dev.to May 16, 2026, 23:10 UTC
EN

SpeakShift: A Fully Local Desktop App Powered by Whisper.cpp + NLLB + FFmpeg

SpeakShift: Fully Local Whisper.cpp + NLLB Translation + FFmpeg Media Converter Hi DEV Community 👋 Like many of you, I spend a lot of time working w…

localllmwhisperaiproductivity
Dev.to May 16, 2026, 06:23 UTC
EN

Choosing the Right Local AI Stack for SOC Alert Triage: Model, Engine, and Harness

Choosing the Right Local AI Stack for SOC Alert Triage: Model, Engine, and Harness Practical guidance for cybersecurity engineers who want local AI…

cybersecurityailocalllmsoc
Dev.to May 16, 2026, 06:21 UTC
EN

Localmaxxing isn't theory. Here's what my 3-GPU rig actually does.

Tom Tunguz wrote a post this week called Localmaxxing . His thesis: open-weight models on prosumer hardware now match cloud-tier quality for a slive…

localllmaieconomicsagentcostcontrolgpuinference
Dev.to May 15, 2026, 14:45 UTC
EN

TextGen vs LM Studio: Picking a Local LLM Runner in 2026

I've been running local LLMs on my workstation for about two years now. Started with llama.cpp raw on the command line, moved to LM Studio when I wa…

localllmaiopensourceproductivity
Dev.to May 14, 2026, 20:17 UTC
EN

Local LLMs in 2026: What Actually Works on Consumer Hardware

Local LLMs in 2026 work on three hardware lanes: 32-core CPU with 64GB+ RAM hits 10-25 tokens per second on Qwen 3 14B, an RTX 4090 hits 30-80 token…

ailocalllmollamaqwen
Dev.to May 10, 2026, 11:36 UTC
EN

[Day 3] I Had a Local LLM Analyze a Year of My Credit Card Statements

[Day 3] I Had a Local LLM Analyze a Year of My Credit Card Statements Intro Day 3: I'm going to hand a year of credit card statements over to a loca…

localllmaidgxsparkollama
Dev.to May 5, 2026, 22:52 UTC
EN

[Day 2] I Trained an AI on 22 Photos of My Cat — Now It Draws Her in Any Scene

[Day 2] I Trained an AI on 22 Photos of My Cat — Now It Draws Her in Any Scene So, yesterday I generated "some cat" Day 1 ended with "I made my DGX…

localllmaidgxsparklora
Dev.to May 5, 2026, 00:06 UTC
EN

[Day 1] DGX Spark Came Home — I Made It Draw a Cat

[Day 1] DGX Spark Came Home — I Made It Draw a Cat So... what is "local LLM" again? Honestly, I'm still figuring out what "local LLM" even means. Bu…

localllmaidgxsparkcomfyui
Dev.to May 4, 2026, 03:20 UTC

© Tech News — Headline Aggregator

English Русский
Sitemap Legal Notice Privacy Terms Copyright / Removal DSA Contact

Leaving the site

You are about to open an external website:

Continue →