Keeps: post title, subreddit, score
<!-- SC_OFF --><div class="md"><p>Hey all, I'm serving DSv4Flash 0731 on a cluster of 2x DGX Sparks but am running into constant issues with having almost no RAM (unified memory) left for the OS/cache and I'd love to hear the community feedback on what I could do to get more RAM
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1vifnby/amazon_cracks_down_on_cpu_waste_among_engineers/"> <img alt="Amazon cracks down on 'CPU waste' among engineers as agentic AI crunch intensifies — CPU demand makes low-utilization EC2 instances a hot comm
reddit:hardware<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1viffzo/sk_hynix_to_invest_54_trillion_won_in/"> <img alt="SK Hynix to Invest 54 Trillion Won in Semiconductor Facilities. New Plants in Yongin, Cheongju to Produce Next-Gen DRAM, NAND by 2029" src="https://exte
reddit:hardware<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1vieca5/nanya_technology_invests_107_billion_in_euv_dram/"> <img alt="Nanya Technology Invests $10.7 Billion in EUV DRAM Expansion" src="https://external-preview.redd.it/nkOxJ9diCKDC3aXj0a_bLU3rO-grJKLy95nAcsW0C
reddit:hardware- 2026-08-07Qwen 3.6 27B flags/settings in llama.cpp
<!-- SC_OFF --><div class="md"><p>I run the following on a 5090 and have been okay with its performance, it does most things somewhere 80-100 t/s, though that can slow down at full 262k context - more like 40 t/s at times. I use it primarily in appdev tasks. This just barely fits
reddit:LocalLLaMA <table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1viamnr/intel_just_matched_apple_silicon_seriously/"> <img alt=" Intel just matched Apple Silicon. Seriously." src="https://external-preview.redd.it/SYaGXS_NgdBhRKjW12BVb8_-L57EUqjRvmbw7QY5iUY.jpeg?width=320&
reddit:hardware<!-- SC_OFF --><div class="md"><p>Hey all,</p> <p>Are there any Delphi developers in the crowd? If so, which models would you say are best at doing development in Delphi? What are your thoughts/suggestions here, and is there a good GUI client/harnass you like for doing delphi spe
reddit:LocalLLaMA- 2026-08-07DeepSeek V4 Flash 0731 - ARC-AGI Results
<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vi9zls/deepseek_v4_flash_0731_arcagi_results/"> <img alt="DeepSeek V4 Flash 0731 - ARC-AGI Results" src="https://external-preview.redd.it/sWhpb1GjRlbd3knWV_xC1C2WMQX5RRFImjvBgSF_7ZI.png?width=640&crop=sma
reddit:LocalLLaMA <!-- SC_OFF --><div class="md"><p>Hey everyone, I just wanted to share my journey here for some motivation.</p> <p>Three years ago, I saw the sudden spike in AI and realized it was the future of tech. My goal at the time was to be an indie game dev, and seeing that AI could write
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1vi7eqk/exclusive_apple_is_scrambling_for_dram_weeks/"> <img alt="Exclusive: Apple is Scrambling for DRAM Weeks Ahead of Ultra, iPhone 18 Launch" src="https://external-preview.redd.it/MMZEuxTmjqXyrLJVbf32Hbik93P
reddit:hardware<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vi77dr/parakeetwgsl_fast_accurate_asr_in_the_browser_via/"> <img alt="parakeet.wgsl – Fast, accurate ASR in the browser, via raw WebGPU & SIMD WASM" src="https://external-preview.redd.it/aGFyZHk4cnZuemhoM
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vi6hmw/llamacpp_pr_reports_up_to_169_faster_quantizedkv/"> <img alt="llama.cpp PR reports up to 169% faster quantized-KV decode at 118K context on Intel Battlemage from one SYCL kernel switch" src="https://pr
reddit:LocalLLaMA- 2026-08-07What happened to the Mobidapter?
<!-- SC_OFF --><div class="md"><p>People used Mobidapters for old phones. It wasn’t a card reader, it was an adapter so you could plug a USB thumb drive into the phone. This could have been really good for Raspberry Pis or Orange Pi 5 Pluses because Microsds are getting more expe
reddit:hardware <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vi1r6t/wananimate2_pushing_the_application_boundaries_of/"> <img alt="Wan-Animate-2: Pushing the Application Boundaries of Character Animation Models" src="https://preview.redd.it/7t00lk39nyhh1.png?width=640&
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vi0d4i/lfm2526b_modelkv_cache_quantization_report/"> <img alt="LFM2.5-2.6B model+KV cache quantization report" src="https://preview.redd.it/3lr47ialdyhh1.png?width=140&height=92&auto=webp&s=e5244a
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vhz989/a_llamacpp_pr_makes_q2_0_3036x_faster_on_x86_cpus/"> <img alt="A llama.cpp PR makes Q2_0 3.0–3.6x faster on x86 CPUs, 8B decode goes 2.39 → 8.20 tok/s" src="https://preview.redd.it/pyim0m155yhh1.jpeg?w
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vhy2e6/rtx_5090_owner_built_an_opensource_tool_that/"> <img alt="RTX 5090 Owner Built An Open-Source Tool That Shuts Down PC If It Detects The 12VHPWR Cable Drawing Too Much Power, But It Can Only Work On Spe
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1vhwqih/price_hikes_may_be_coming_for_pc_motherboards_next/"> <img alt="Price Hikes May Be Coming for PC Motherboards Next" src="https://external-preview.redd.it/OuRNAj98ocMRlkcLHuwjKZ6me96KJ_CF95mgg_z8UDI.jpeg?
reddit:hardware<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vhwilp/an_openweight_model_too_moonshot_joins_the_race/"> <img alt="An open-weight model too, Moonshot joins the race (gently this time)" src="https://preview.redd.it/6i806mqxexhh1.jpeg?width=640&crop=sma
reddit:LocalLLaMA<!-- SC_OFF --><div class="md"><p>Hello,</p> <p>For the past few days I have been benchmarking Gemma 4 26b QAT UD Q4_K_XL extensively versus Bartowski's Q4_K_L.</p> <p>While QAT is certainly very effective and reducing memory consumption versus the highest q4 quant from him, I al
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vhv2bz/ds4_flash_incoming_price_increase_weve_been_able/"> <img alt="DS4 Flash incoming price increase "we've been able to reproduce their current prices even on rented GPUs"" src="https://preview.r
reddit:LocalLLaMA<!-- SC_OFF --><div class="md"><p>I'm using AWS SAM to provision API Gateway, Lambdas, and RDS. I want to know how I can do database migrations easily without having to keep up an EC2 instance just for the purpose of connecting to the RDS db and running db migrations?</p> <p>I wa
reddit:devops<table> <tr><td> <a href="https://www.reddit.com/r/NVDA_Stock/comments/1vhrm3o/semianalysis_suggests_gemini_subsequent_models/"> <img alt="SemiAnalysis suggests Gemini subsequent models will not be competitive because top researchers prefer Nvidia GPUs over TPUs and are willing t
reddit:NVDA_Stock<!-- SC_OFF --><div class="md"><p>This post contains content not supported on old Reddit. <a href="https://sh.reddit.com/r/NVDA_Stock/comments/1vhq2e9">Click here to view the full post</a></p> </div><!-- SC_ON -->   submitted by   <a href="https://www.reddit.com/user/dail
reddit:NVDA_Stock- 2026-08-07AMD acquires AI chip startup Taalas to boost inference performance by etching models into silicon
  submitted by   <a href="https://www.reddit.com/user/SirActionhaHAA"> /u/SirActionhaHAA </a> <br /> <span><a href="https://www.theregister.com/systems/2026/08/06/amd-acquires-ai-chip-startup-taalas-to-boost-inference-performance-by-etching-models-into-silicon/5284344">[l
reddit:hardware <table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1vhmmrr/exclusive_this_may_be_lenovos_thinnest_thinkbook/"> <img alt="Exclusive: This may be Lenovo’s thinnest ThinkBook laptop yet" src="https://external-preview.redd.it/oDFRSNN7gnGBfkFyV4BABmxwWRCLsRxxhfOWyZea
reddit:hardware- 2026-08-07Thoughts on ADO?
<!-- SC_OFF --><div class="md"><p>I'm being brought in to assist with getting a devops plan in place. The current team is using ADO for issue management and code repos. They have no real existing IaC, CI-CD, monitoring, etc.</p> <p>My first reaction is to tell them to run from AD
reddit:devops