MINMAER

AI Tools & Local AI

SHEET · UPDATED 2026-09-28 · SORT ANY COLUMN

Start with a $10 coding plan or metered API. For local models, buy enough memory first, then compare measured decode speed per dollar.

Sheet a · AI Tools & Local AI · New in AI (Aug to Sep 2026)Updated 2026-09-28
AI Tools & Local AI · New in AI (Aug to Sep 2026): specs, price and value score per product
Link
GoogleGemini 3.8 Flash: reasoning and coding with a free developer tierTOP VALUE
Launched 2026-09-02; GA; gemini-3.8-flash 1M context; multimodal input and tool use Access: Google AI Studio / Gemini API free tier; quotas and regional eligibility apply Editorial value/novelty 90/100; strong agent capabilities at a low entry cost; vendor benchmark claims Free developer usage is limited. Gemini app access requires AI Pro or Ultra. Paid API: $0.75 input / $3.75 output per 1M through 2026-12-31; then $1.50 / $7.50. 90 FreeFree/pt 100 Check price →
Alibaba / QwenQwen3.8-27B: open multimodal model for local coding and work
August 2026 release; listed in QwenCloud 2026-08-19; GA open weights 27B dense; 262,144 native context, extensible to 1M Access: Qwen/Qwen3.8-27B on Hugging Face; Transformers, vLLM or compatible quantized runtime Editorial value/novelty 89/100; useful local size and no model subscription; published results are vendor evaluations Apache 2.0. Zero is the software price on hardware you already own; memory, electricity and hosted inference cost extra. Qwen4 is not substituted for released weights. 89 FreeFree/pt 99 Check price →
TypeSafe AIJev: typed decision model for routing, scoring and automation
Launched 2026-09-15; early access; current documented model Jev 1.13 64K request budget; state plus longest question limited to 32K; text input Access: request early access at typesafe.ai, then use console.typesafe.ai and the SDK/API Editorial value/novelty 88/100; unusual low-cost decision primitive; speed comparisons are vendor tests Zero denotes joining early access, not free inference. Published API price: $0.042 per 1M input tokens; output free. No fixed monthly plan verified. Returns choices, scores and yes/no judgments, not prose; decisions can still be wrong. 88 FreeFree/pt 98 Check price →
MetaMuse: personal agent that carries out tasks in its own cloud computer
Launched 2026-09-08; beta / limited testing; powered by Muse Spark Dedicated VM, browser and connected apps; continues work after you leave Access: muse.ai or Muse iOS/Android, with WhatsApp messaging; regional and age eligibility apply Editorial value/novelty 86/100; broad free task execution; early users report both useful results and repeated omissions Free usage-limited tier. Power $20/month: 500M Muse tokens/week; Maximum $100/month: 3B/week. US launch; check local onboarding. Purchases made by the agent are extra. 86 FreeFree/pt 96 Check price →
Epoch AIBenchmark hub + Epoch Brief: independent tracking of AI progress
GA; September refresh: Epoch Brief 2026-09-12 and benchmark analysis 2026-09-16 Capabilities index, math/coding evaluations, compute and cost research Access: read epoch.ai and subscribe to the free Epoch Brief at epochai.substack.com Editorial value/novelty 86/100; useful free evidence for choosing models and checking launch claims This means the Epoch AI research organization and newsletter, not Sarvam. Sarvam Epoch is an event featuring Sarvam 105B updates, not a verified model named Epoch. The hub predates this window; its new reports are the inclusion. 86 FreeFree/pt 96 Check price →
OpenAICodex: free Luna desktop entry and September agent tooling updates
Update 2026-09-22; GA with staged rollout; GPT-6 Luna on Free Desktop coding tasks; model and task size determine available usage Access: ChatGPT Free account in the desktop app where Luna rollout is enabled Editorial value/novelty 85/100; zero-cost entry to repository work; September CLI also adds marketplace plugin management Free route is limited desktop Luna use. Plus $20/month adds Sol and paid-plan surfaces including CLI, IDE and cloud integrations. Do not assume every Codex feature is included on Free. 85 FreeFree/pt 94 Check price →
SpaceXAIGrok 4.7: coding and knowledge-work model with free Grok Build entry
Launched 2026-09-21; GA Long-running coding, documents and presentations; standard and faster serving variants Access: try Grok Build at x.ai/build; also available through Cursor and the Grok API Editorial value/novelty 84/100; accessible new coding option; claimed speed and quality gains are vendor comparisons Launch page offers free Grok Build use, subject to limits. API is separate: starts at $2 input / $6 output per 1M tokens; fast variant costs twice as much. Latest verified release is 4.7, not 5. 84 FreeFree/pt 93 Check price →
DeepSeekDeepSeek-V4.1-Flash: open multimodal model with a cheaper agent API
Launched 2026-09-10; GA; API model ID deepseek-flash 552B backbone MoE; 8B active input / 16B output; 1M context Access: MIT weights at deepseek-ai/DeepSeek-V4.1-Flash on Hugging Face; metered hosted API available Editorial value/novelty 84/100; lower serving costs and smaller KV cache; benchmark gains are provider-reported Zero is the open-weight software subscription price, not free hosted API or hardware. This is a large deployment, not an 8B desktop model. Most enthusiasts should compare metered API costs before attempting local hosting. 84 FreeFree/pt 93 Check price →
OllamaClaude Desktop gateway: run open models through the Claude interface
Integration launched 2026-08-25; GA Local models on your computer or separately hosted Ollama models Access: install Ollama and Claude Desktop, enable Claude in Ollama settings, then choose a local model Editorial value/novelty 83/100; makes local models usable through a familiar agent interface without a model subscription Zero applies to the local-model route with existing hardware. Cloud plans and compute cost extra. Toggling the gateway back restores the previous Claude setup; Anthropic model access remains separate. 83 FreeFree/pt 92 Check price →
MetaMuse Glimmer 30B: open multimodal model for local personal agents
Public release documented 2026-08-10; GA open weights 30B dense; approximately 131K context; text and image understanding Access: meta-models/Muse-Glimmer-30B on Hugging Face; GGUF, Transformers, llama.cpp and vLLM support Editorial value/novelty 81/100; useful local-agent release, but early hands-on coding reports are mixed; Qwen remains the stronger value pick here Apache 2.0; zero software subscription, hardware and electricity extra. This is Meta's verified new open release in the window; no new Llama-branded release was verified. Separate from the hosted Muse personal agent. 81 FreeFree/pt 90 Check price →
Mistral AIShieldstral: small open model for custom text and image safety checks
Launched 2026-08-04; GA open weights 3B parameters; provider documents operation on one 16GB NVIDIA GPU Access: Hugging Face weights linked from Mistral's Shieldstral announcement Editorial value/novelty 80/100; inexpensive specialized classification; comparison with larger guard models is a vendor evaluation Apache 2.0. Provide a policy as a question and receive a safety score without retraining. Zero software subscription with your own hardware; not a general chat model. 80 FreeFree/pt 89 Check price →
AppleSiri AI: personal-context assistant with onscreen awareness and app actions
Rollout launched 2026-09-14; beta in English Messages, mail, photos, screen context and supported systemwide app actions Access: Apple's 2027 OS releases on compatible Apple Intelligence hardware; enable Siri AI where available Editorial value/novelty 79/100; useful integration at no added subscription cost for existing eligible-device owners Compatible hardware is required and costs extra if you do not own it. Languages, regions and devices limit access. Listed as a shipped beta, not an announced future assistant. 79 FreeFree/pt 88 Check price →
Spear Street TechnologyInstinct: invite-only personal agent over iMessage and WhatsApp
Invite beta active by August 2026; exact first-launch day not publicly verified Personal tasks through messaging, calls and connected applications Access: instinct.com / app.instinct.com; request an existing member invite or join the waitlist Editorial value/novelty 72/100; novel personal-task automation, discounted for scarce access and limited independent evidence No public subscription price verified; zero denotes unpriced invite access, not guaranteed permanent free service. Account permissions and action reliability need hands-on evaluation. 72 FreeFree/pt 80 Check price →
OpenAIGPT-6 Sol: lower-cost model for complex coding and agent tasks
Codex / ChatGPT Work rollout 2026-09-22; GA with staged account availability API: 1,050,000 context and 128,000 maximum output; app limits vary Access: ChatGPT Plus, select Sol in Codex or Work; API is billed separately Editorial value/novelty 89/100; better sustained-work value than Astra when Sol solves the task $20/month billed monthly includes Codex allowance. Free and Go receive Luna, not Sol. Sol uses fewer credits than Astra; model access and quotas depend on rollout. 89 $20.00$0.22/pt 5 Check price →
AnthropicClaude Code: September agent updates and the new Opus 5.5 default
Update 2026-09-22, v2.1.280; GA; Opus 5.5 becomes the default Opus model Opus 5.5 supports 1M context; repository, terminal and tool workflows Access: install/update Claude Code and sign in with Claude Pro; subsequent September fixes are available Editorial value/novelty 88/100; useful coding upgrade within Pro; release notes also fix symlink approval and retry behavior $20/month billed monthly; shares plan limits with Claude. Extra usage and API billing are separate. This covers the Opus update, not included Fable usage on Pro. 88 $20.00$0.23/pt 5 Check price →
OpenAIGPT-6 Astra: frontier reasoning for difficult coding and research
September 2026 release; Codex picker support verified 2026-09-04; GA API: 1,050,000 context and 128,000 maximum output; app limits vary Access: ChatGPT Plus via Codex / Work model picker; separate paid API also available Editorial value/novelty 87/100; Epoch finds exceptional math capability, but coding gains depend on task and cost $20/month billed monthly provides limited included use. Astra consumes substantially more allowance than Sol; choose it for difficult tasks that justify the cost. 87 $20.00$0.23/pt 5 Check price →
AnthropicClaude Cowork: tasks, documents and chat in one conversation
Update 2026-09-16; GA staged rollout to Pro and Max Reports, spreadsheets and presentations with carried-over projects, skills and connectors Access: Claude Pro on web, desktop or mobile as the unified experience reaches the account Editorial value/novelty 85/100; fewer mode switches for everyday knowledge work; not a new underlying model $20/month billed monthly for Pro; shared usage limits. Fable 5.1 requires separately paid credits on Pro. Some artifact features are free, but that is not the full Cowork task workflow. 85 $20.00$0.24/pt 5 Check price →
CursorProjects: a coordinator for persistent, multi-agent coding work
Launched 2026-09-10; beta rollout; self-hosted machines added 2026-09-02 Shared project context across agents; scheduled work and event subscriptions Access: Projects in Cursor; Pro is the verified monthly plan with cloud-agent execution Editorial value/novelty 82/100; useful coordination for larger jobs, with spend increasing across parallel agents $20/month billed monthly. Projects UI is rolling out to all users, but free Hobby access does not establish included cloud execution. Sept 23 Rollouts and Security Review require Teams or Enterprise. 82 $20.00$0.24/pt 5 Check price →
CognitionDevin Code Scans: investigate a codebase and turn findings into PRs
Launched 2026-09-16; GA; macOS cloud support announced 2026-09-15 Parallel investigation, combined findings and follow-up code changes Access: Devin web app, enter /scan; Pro includes Devin Cloud Editorial value/novelty 81/100; useful for broad cleanup and performance tasks; published time savings are selected vendor/customer cases $20/month billed monthly with limited quota; extra usage costs more. Free desktop/CLI access does not include the cloud workflow listed here. macOS access depends on workspace availability. 81 $20.00$0.25/pt 5 Check price →
AnthropicClaude Fable 5.1: advanced coding and research, with metered Pro access
Launched 2026-09-01; GA; claude-fable-5-1 1M context; 128K maximum API output; long-running coding and knowledge work Access: Claude Pro with paid usage credits, or Max with included Fable allowance; separate API available Editorial value/novelty 78/100; high capability, but extra charges reduce entry-level value versus included Opus 5.5 $20 is the monthly Pro base, not the all-in cost: Fable requires additional metered credits from the first use. Max starts at $100/month with up to 50% of its weekly allowance usable on Fable. No flat cheapest all-in monthly total exists. 78 $20.00$0.26/pt 4 Check price →
Sheet note: New in AI: prices are USD per month for the stated entry route, using monthly billing. Zero can mean a limited free tier, free local software on existing hardware, or unpriced invite/early-access signup; read each row. Jev API calls are paid, and Fable requires paid credits above its Pro base. New-section scores include usefulness per dollar, novelty and access friction, with brand recognition at zero weight. They are provisional editorial synthesis, not a measured community poll. Dates distinguish launches from updates and report editions. Original four sections, preserved unchanged, use the following methodology: Prices mix USD/month subscriptions, USD per 1M uncached input tokens for APIs, and one-time hardware. API output rates are in Price basis; API access has no included monthly allowance. Compare value only within the same kind and workload: the site-wide perf/price index cannot rank these billing units together. Listed API ordering assumes 1M input plus 1M output; other kinds use perf/price within kind. GPU prices exclude a host PC; Framework is a refurbished barebone, while other boxes are complete. Scores are editorial, brand recognition has zero weight, and price is not baked into perf. Hardware scores combine reported 7B Q4 decode with model-fitting memory; missing measurements use memory/bandwidth proxies. Cross-build benchmarks are indicative, not controlled comparisons. Subscription and new-model scores are provisional usefulness judgments, not a measured consensus poll. Used and marketplace offers can expire; SOURCES.md records evidence and limitations. GPU and box prices are the in-stock price on the check date; the 2026 memory squeeze keeps them volatile, and used 3090 asks are far above their 2025 level, so treat the hardware sections as a snapshot.

MMX score: editorial bench, 0 to 100, per channel. Value index: points per dollar, scaled so the channel's best ratio = 100.

Sheet b · AI Tools & Local AI · Subscriptions (USD per month)Updated 2026-09-28
AI Tools & Local AI · Subscriptions (USD per month): specs, price and value score per product
Link
GitHubCopilot ProTOP VALUE
Copilot IDE, CLI, agent and code review; model availability varies Context depends on selected model and tool $15 AI credits currently: $10 base + $5 variable flex Editorial usefulness 77/100; budget coding assistance Credit-metered agent usage; do not assume legacy premium-request quotas. 77 $10.00$0.13/pt 100 Check price →
OpenAIChatGPT Plus with Codex
ChatGPT and Codex; GPT-6 Sol, Luna and Astra access in Codex App context depends on model and surface Shared usage allowances; task size, model and reasoning change consumption Editorial usefulness 88/100; general assistant plus coding agent Codex is included, not a second $20 subscription. API usage is separate. 88 $20.00$0.23/pt 57 Check price →
AnthropicClaude Pro with Claude Code
Claude assistant, Claude Code and Cowork; models depend on plan App context depends on selected model Claude and Claude Code share plan limits; extra usage can cost more Editorial usefulness 87/100; writing and repository work Monthly price; annual billing has a different effective rate. API usage is separate. 87 $20.00$0.23/pt 56 Check price →
GoogleGoogle AI Pro
Gemini app, Gemini 3.1 Pro, Jules and expanded Antigravity access Up to 1M-token app context Higher plan limits; each service has its own quota Editorial usefulness 84/100; long-context and Google workflows Includes 5TB storage; API tokens are billed separately. Regular monthly price, no trial. 84 $19.99$0.24/pt 55 Check price →
CursorCursor Pro
Multi-model IDE agents, frontier models, cloud agents and MCP Context and model availability vary Extended agent usage; no fixed universal request count Editorial usefulness 83/100; integrated coding workflow Additional usage and Bugbot billing can increase cost. 83 $20.00$0.24/pt 54 Check price →
CognitionDevin Pro (Windsurf pricing successor)
Devin/Windsurf editor; OpenAI, Claude, Gemini and open-model choices Context depends on selected model Increased usage quota; unlimited Tab and inline edits Editorial usefulness 80/100; editor and agent workflow Windsurf pricing redirected to Devin pricing when checked; do not reuse the old $15 plan. 80 $20.00$0.25/pt 52 Check price →
PerplexityPerplexity Pro
Cited search with OpenAI, Anthropic and Google model options Search context is service-managed Expanded research and model access; published limits can change Editorial usefulness 78/100; research with source discovery Consumer subscription; Sonar API spend is separate. 78 $20.00$0.26/pt 51 Check price →
KagiUltimate
Kagi Search and Assistant with flagship models and Research Context varies by Assistant model Unlimited search; Assistant has fair-use spending limits Editorial usefulness 81/100; private search plus multi-model assistant Assistant is not unlimited API credit; additional payments may be needed. 81 $25.00$0.31/pt 42 Check price →
AnthropicClaude Max 5x
Claude assistant and Claude Code App context depends on selected model 5x Pro usage; shared Claude/Code limits Editorial usefulness 90/100; higher sustained agent allowance Same ecosystem as Pro; worth comparing only if Pro quotas constrain work. 90 $100.00$1.11/pt 12 Check price →
OpenAIChatGPT Pro 5x with Codex
ChatGPT and Codex; GPT-6 model family App context depends on selected model 5x usage tier; model and task affect consumption Editorial usefulness 90/100; higher sustained agent allowance 5x tier, not the $200 20x tier. API usage is separate. 90 $100.00$1.11/pt 12 Check price →
Sheet note: New in AI: prices are USD per month for the stated entry route, using monthly billing. Zero can mean a limited free tier, free local software on existing hardware, or unpriced invite/early-access signup; read each row. Jev API calls are paid, and Fable requires paid credits above its Pro base. New-section scores include usefulness per dollar, novelty and access friction, with brand recognition at zero weight. They are provisional editorial synthesis, not a measured community poll. Dates distinguish launches from updates and report editions. Original four sections, preserved unchanged, use the following methodology: Prices mix USD/month subscriptions, USD per 1M uncached input tokens for APIs, and one-time hardware. API output rates are in Price basis; API access has no included monthly allowance. Compare value only within the same kind and workload: the site-wide perf/price index cannot rank these billing units together. Listed API ordering assumes 1M input plus 1M output; other kinds use perf/price within kind. GPU prices exclude a host PC; Framework is a refurbished barebone, while other boxes are complete. Scores are editorial, brand recognition has zero weight, and price is not baked into perf. Hardware scores combine reported 7B Q4 decode with model-fitting memory; missing measurements use memory/bandwidth proxies. Cross-build benchmarks are indicative, not controlled comparisons. Subscription and new-model scores are provisional usefulness judgments, not a measured consensus poll. Used and marketplace offers can expire; SOURCES.md records evidence and limitations. GPU and box prices are the in-stock price on the check date; the 2026 memory squeeze keeps them volatile, and used 3090 asks are far above their 2025 level, so treat the hardware sections as a snapshot.

MMX score: editorial bench, 0 to 100, per channel. Value index: points per dollar, scaled so the channel's best ratio = 100.

Sheet c · AI Tools & Local AI · API pricing (USD per 1M input tokens)Updated 2026-09-28
AI Tools & Local AI · API pricing (USD per 1M input tokens): specs, price and value score per product
Link
OpenAIGPT-6 Luna APITOP VALUE
GPT-6 Luna API 1,050,000 context; 128,000 max output Standard uncached short-context rates; account-tier rate limits Editorial capability score; not a benchmark percentage Long-context rates: $0.20 input / $0.75 output per 1M. No monthly subscription included. 73 $0.10$0.00/pt 100 Check price →
DeepSeekDeepSeek-V4.1-Flash API
DeepSeek-V4.1-Flash API 1M context; 384K max output Peak uncached rate; published concurrency 2500 Editorial capability score; not a benchmark percentage Model ID deepseek-flash; off-peak $0.15 input / $0.60 output per 1M. 81 $0.30$0.00/pt 37 Check price →
GoogleGemini 3.8 Flash API
Gemini 3.8 Flash API 1,048,576 input; 65,536 max output Paid text/image/video rates; account-tier quotas Editorial capability score; not a benchmark percentage Promotional through 2026-12-31; then $1.50 input / $7.50 output per 1M. Audio has separate pricing. 86 $0.75$0.01/pt 16 Check price →
DeepSeekDeepSeek-V4-Pro-0813 API
DeepSeek-V4-Pro-0813 API 1M context; 384K max output Peak uncached rate; published concurrency 500 Editorial capability score; not a benchmark percentage Model ID deepseek-v4-pro; off-peak $0.66 input / $1.98 output per 1M. 85 $1.32$0.02/pt 9 Check price →
OpenAIGPT-6 Sol API
GPT-6 Sol API 1,050,000 context; 128,000 max output Standard uncached short-context rates; account-tier rate limits Editorial capability score; not a benchmark percentage Long-context rates: $4 input / $15 output per 1M. 87 $2.00$0.02/pt 6 Check price →
AnthropicClaude Sonnet 5 API
Claude Sonnet 5 API 1M context; 128K max output Standard uncached rates; account-tier limits Editorial capability score; not a benchmark percentage Model ID claude-sonnet-5; API billing is separate from Claude plans. 86 $2.00$0.02/pt 6 Check price →
AnthropicClaude Opus 5.5 API
Claude Opus 5.5 API 1M context; 128K max output Standard uncached rates; account-tier limits Editorial capability score; not a benchmark percentage Model ID claude-opus-5-5; cache and batch discounts excluded. 89 $4.00$0.04/pt 3 Check price →
OpenAIGPT-6 Astra API
GPT-6 Astra API 1,050,000 context; 128,000 max output Standard uncached short-context rates; account-tier rate limits Editorial capability score; not a benchmark percentage Long-context rates: $20 input / $75 output per 1M. 90 $10.00$0.11/pt 1 Check price →
AnthropicClaude Fable 5.1 API
Claude Fable 5.1 API 1M context; 128K max output Standard uncached rates; account-tier limits Editorial capability score; not a benchmark percentage Model ID claude-fable-5-1; recent release, limited independent longitudinal evidence. 90 $10.00$0.11/pt 1 Check price →
Sheet note: New in AI: prices are USD per month for the stated entry route, using monthly billing. Zero can mean a limited free tier, free local software on existing hardware, or unpriced invite/early-access signup; read each row. Jev API calls are paid, and Fable requires paid credits above its Pro base. New-section scores include usefulness per dollar, novelty and access friction, with brand recognition at zero weight. They are provisional editorial synthesis, not a measured community poll. Dates distinguish launches from updates and report editions. Original four sections, preserved unchanged, use the following methodology: Prices mix USD/month subscriptions, USD per 1M uncached input tokens for APIs, and one-time hardware. API output rates are in Price basis; API access has no included monthly allowance. Compare value only within the same kind and workload: the site-wide perf/price index cannot rank these billing units together. Listed API ordering assumes 1M input plus 1M output; other kinds use perf/price within kind. GPU prices exclude a host PC; Framework is a refurbished barebone, while other boxes are complete. Scores are editorial, brand recognition has zero weight, and price is not baked into perf. Hardware scores combine reported 7B Q4 decode with model-fitting memory; missing measurements use memory/bandwidth proxies. Cross-build benchmarks are indicative, not controlled comparisons. Subscription and new-model scores are provisional usefulness judgments, not a measured consensus poll. Used and marketplace offers can expire; SOURCES.md records evidence and limitations. GPU and box prices are the in-stock price on the check date; the 2026 memory squeeze keeps them volatile, and used 3090 asks are far above their 2025 level, so treat the hardware sections as a snapshot.

MMX score: editorial bench, 0 to 100, per channel. Value index: points per dollar, scaled so the channel's best ratio = 100.

Sheet d · AI Tools & Local AI · Local AI GPUs (VRAM per dollar)Updated 2026-09-28
AI Tools & Local AI · Local AI GPUs (VRAM per dollar): specs, price and value score per product
Link
ASRock / IntelArc B580 Challenger 12GBTOP VALUE
Local GGUF models through Vulkan; Intel backends vary 12GB dedicated VRAM; 456 GB/s nominal bandwidth 7B Q4 fits; larger models and long KV caches need more memory 70.14 tokens/s: Llama 2 7B Q4_0, Vulkan tg128, no FA New; requires host PC and compatible software. GPU-family benchmark, not this exact board. 54 $309.99$5.74/pt 100 Check price →
MSI / NVIDIARTX 5060 Ti SHADOW 2X OC PLUS 16GB
CUDA and Vulkan local models 16GB dedicated VRAM; 448 GB/s nominal bandwidth 7B to 14B Q4 practical; context consumes additional VRAM 93.51 tokens/s: Llama 2 7B Q4_0, Vulkan tg128, no FA New, sold and shipped by Newegg; GPU-family benchmark does not establish the tested board or capacity. 66 $779.99$11.82/pt 49 Check price →
GIGABYTE / AMDRadeon RX 7900 XTX GAMING OC 24GB
Vulkan local models; ROCm support depends on OS and runtime 24GB dedicated VRAM; 960 GB/s nominal bandwidth More model and KV-cache headroom than 16GB cards 182.63 tokens/s: Llama 2 7B Q4_0, Vulkan tg128, no FA New marketplace listing, roboshine, Hong Kong; GPU-family result, not exact retail board. 90 $1309.34$14.55/pt 39 Check price →
MSI / NVIDIARTX 4060 Ti GAMING X SLIM 16GB
CUDA and Vulkan local models 16GB dedicated VRAM; 288 GB/s nominal bandwidth 7B to 14B Q4 practical; memory bandwidth limits decode No matching benchmark verified; score uses VRAM and bandwidth proxy New marketplace listing, TECH EDGE, ships from China. Poor value at this quote versus the 5060 Ti. 57 $865.00$15.18/pt 38 Check price →
GIGABYTE / NVIDIAGeForce RTX 3090 24GB (used)
CUDA and Vulkan local models 24GB dedicated VRAM; 936 GB/s nominal bandwidth Used-card condition, cooling and host PSU matter 164.05 tokens/s: Llama 2 7B Q4_0, Vulkan tg128, no FA Single used listing; $40 shipping extra. $1000 opening bid is not the purchase price. Weak value at $2000. 86 $2000.00$23.26/pt 25 Check price →
Sheet note: New in AI: prices are USD per month for the stated entry route, using monthly billing. Zero can mean a limited free tier, free local software on existing hardware, or unpriced invite/early-access signup; read each row. Jev API calls are paid, and Fable requires paid credits above its Pro base. New-section scores include usefulness per dollar, novelty and access friction, with brand recognition at zero weight. They are provisional editorial synthesis, not a measured community poll. Dates distinguish launches from updates and report editions. Original four sections, preserved unchanged, use the following methodology: Prices mix USD/month subscriptions, USD per 1M uncached input tokens for APIs, and one-time hardware. API output rates are in Price basis; API access has no included monthly allowance. Compare value only within the same kind and workload: the site-wide perf/price index cannot rank these billing units together. Listed API ordering assumes 1M input plus 1M output; other kinds use perf/price within kind. GPU prices exclude a host PC; Framework is a refurbished barebone, while other boxes are complete. Scores are editorial, brand recognition has zero weight, and price is not baked into perf. Hardware scores combine reported 7B Q4 decode with model-fitting memory; missing measurements use memory/bandwidth proxies. Cross-build benchmarks are indicative, not controlled comparisons. Subscription and new-model scores are provisional usefulness judgments, not a measured consensus poll. Used and marketplace offers can expire; SOURCES.md records evidence and limitations. GPU and box prices are the in-stock price on the check date; the 2026 memory squeeze keeps them volatile, and used 3090 asks are far above their 2025 level, so treat the hardware sections as a snapshot.

MMX score: editorial bench, 0 to 100, per channel. Value index: points per dollar, scaled so the channel's best ratio = 100.

Sheet e · AI Tools & Local AI · Local AI boxes (unified memory)Updated 2026-09-28
AI Tools & Local AI · Local AI boxes (unified memory): specs, price and value score per product
Link
GMKtecNucBox K13 Core Ultra 7 256V, 16GB/512GBTOP VALUE
Arc 140V iGPU; Intel NPU-compatible local runtimes 16GB shared LPDDR5X-8533; 512GB SSD Shared memory includes OS; NPU needs supported models and runtime No matching 7B Q4 result; memory proxy; NPU rated 47 INT8 TOPS Complete US-plug system. TOPS is not tokens/s; RAM is not upgradeable. 43 $589.99$13.72/pt 100 Check price →
FrameworkDesktop DIY Max+ 395, 64GB (refurbished barebone)
Ryzen AI Max+ 395; Radeon 8060S; Vulkan/ROCm as supported 64GB shared LPDDR5X-8000; 256 GB/s theoretical No SSD, OS, AC cable, CPU fan, tiles or expansion cards included No exact configuration benchmark verified; memory and bandwidth proxy Refurbished DIY base, not a complete PC; add missing parts before comparing total cost. 79 $1659.00$21.00/pt 65 Check price →
GMKtecEVO-X2 Ryzen AI Max+ 395, 64GB/1TB
Ryzen AI Max+ 395; Radeon 8060S; Windows 11 Pro 64GB shared LPDDR5X-8000; 256 GB/s theoretical; 1TB SSD Unified memory includes OS; GPU allocation and backend affect fit No exact configuration benchmark verified; memory and bandwidth proxy US-plug 64GB model at this price, not 128GB. 50-TOPS NPU is not a decode benchmark. 79 $2199.99$27.85/pt 49 Check price →
AppleMac mini M4 Pro 12C/16C, 48GB/2TB (refurbished)
M4 Pro; Metal and MLX local-model runtimes 48GB unified memory; 273 GB/s; 2TB SSD OS and GPU memory allocation reduce usable model capacity No matching 7B Q4 result verified; memory and bandwidth proxy Apple-certified refurbished configuration G1JVCLL/A. Do not compare this 2TB bundle with base-storage quotes. 75 $2549.00$33.99/pt 40 Check price →
NVIDIADGX Spark GB10, 128GB/4TB
GB10 Grace Blackwell; CUDA on Linux ARM64 128GB unified memory; 273 GB/s; 4TB SSD Large-model capacity; ARM64 software compatibility needs checking 58.51 tokens/s: GB10, Llama 2 7B Q4_0, Vulkan tg128, no FA Score reflects 128GB model capacity, not a claim of fastest 7B decode. 90 $4699.00$52.21/pt 26 Check price →
Sheet note: New in AI: prices are USD per month for the stated entry route, using monthly billing. Zero can mean a limited free tier, free local software on existing hardware, or unpriced invite/early-access signup; read each row. Jev API calls are paid, and Fable requires paid credits above its Pro base. New-section scores include usefulness per dollar, novelty and access friction, with brand recognition at zero weight. They are provisional editorial synthesis, not a measured community poll. Dates distinguish launches from updates and report editions. Original four sections, preserved unchanged, use the following methodology: Prices mix USD/month subscriptions, USD per 1M uncached input tokens for APIs, and one-time hardware. API output rates are in Price basis; API access has no included monthly allowance. Compare value only within the same kind and workload: the site-wide perf/price index cannot rank these billing units together. Listed API ordering assumes 1M input plus 1M output; other kinds use perf/price within kind. GPU prices exclude a host PC; Framework is a refurbished barebone, while other boxes are complete. Scores are editorial, brand recognition has zero weight, and price is not baked into perf. Hardware scores combine reported 7B Q4 decode with model-fitting memory; missing measurements use memory/bandwidth proxies. Cross-build benchmarks are indicative, not controlled comparisons. Subscription and new-model scores are provisional usefulness judgments, not a measured consensus poll. Used and marketplace offers can expire; SOURCES.md records evidence and limitations. GPU and box prices are the in-stock price on the check date; the 2026 memory squeeze keeps them volatile, and used 3090 asks are far above their 2025 level, so treat the hardware sections as a snapshot.

MMX score: editorial bench, 0 to 100, per channel. Value index: points per dollar, scaled so the channel's best ratio = 100.

Channel wire

Curated AI Tools & Local AI stories, filtered by MinMaxxer. Snapshot as of 2026-09-28 — refreshed by our editors, not live.