Keeps: post title, subreddit, score
<!-- SC_OFF --><div class="md"><p>So I'm planning to start a Fin-tech app, with financial database and equity reports about all the listed companies in my country. It will have other features as well of course. There are 2 things tho-</p> <ol> <li>I am 14</li> <li>I have a budget
reddit:startups  submitted by   <a href="https://www.reddit.com/user/imaginary_num6er"> /u/imaginary_num6er </a> <br /> <span><a href="https://www.guru3d.com/story/sandisk-launches-520-and-320-sata-ssds-with-capacities-up-to-4tb/">[link]</a></span>   <span><a href="https://www.reddi
reddit:hardware<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1upu4ri/rx_7800_xt_vs_rtx_4070_in_2026_now_with_fsr_4_the/"> <img alt="RX 7800 XT vs RTX 4070 in 2026 (Now with FSR 4!)- The Ultimate Comparison" src="https://external-preview.redd.it/FEipYlZlMyVd7AwI_IJ-lmeOSv8
reddit:hardware<!-- SC_OFF --><div class="md"><p>I have tried YC, Wellfound, and Work at a Startup, but that didn't work</p> <p>After that, I tried reaching out to startups working in worldbuilding and scriptwriting tools, since I’m most interested in that niche. found their discords and emails
reddit:startups<!-- SC_OFF --><div class="md"><p>Is it only me? I'm getting fairly often (every ~2 weeks) emails of people claiming to be "investors" or "M&A advisors" looking to invest or acquire my startup.</p> <p>At the beginning I just disregarded them because they w
reddit:startups<!-- SC_OFF --><div class="md"><p>What's your ultimate goal as a software developer?</p> <p>Is it:</p> <p>- Working in your dream country</p> <p>- Earning a high salary</p> <p>- Great work-life balance</p> <p>- Lower taxes and better quality of life</p> <p>- Working at a top tech
reddit:ExperiencedDevs<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1uplvrp/nvidias_nextgen_ai_rack_system_kyber_nvl144/"> <img alt="Nvidia’s next-gen AI rack system (Kyber NVL144) delayed to 2028 on manufacturing snags, SemiAnalysis says, but Nvidia denies" src="https://externa
reddit:hardware- 2026-07-07Experienced Devs Weekly Burnout and Venting Thread: A weekly thread for sharing experiences
<!-- SC_OFF --><div class="md"><p>This thread is specifically for venting / sharing experiences related to burn-out or similar issues that experienced devs face.</p> </div><!-- SC_ON -->   submitted by   <a href="https://www.reddit.com/user/AutoModerator"> /u/AutoModerato
reddit:ExperiencedDevs   submitted by   <a href="https://www.reddit.com/user/NamelessVegetable"> /u/NamelessVegetable </a> <br /> <span><a href="https://www.theregister.com/systems/2026/07/07/ibm-teases-new-rackable-mainframes-that-complete-the-z17-family/5267423">[link]</a></span>   <span>
reddit:hardware<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1upea2o/china_memory_module_giants_firsthalf_profit_set/"> <img alt="China memory module giant’s first-half profit set to jump more than 600-fold" src="https://external-preview.redd.it/Gl0u1wDk2HI1LSxrvwvF48266D
reddit:hardware- 2026-07-06VideoCardz: "MSI claims first DDR5-8000+ validation for Chinese CXMT memory on AMD motherboards"
<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1updvsn/videocardz_msi_claims_first_ddr58000_validation/"> <img alt="VideoCardz: "MSI claims first DDR5-8000+ validation for Chinese CXMT memory on AMD motherboards"" src="https://external-preview.redd
reddit:hardware <!-- SC_OFF --><div class="md"><p>I'm hoping this doesn't fall inside the "general career advice" rule, otherwise apologies, just remove this post! I'm wondering about the long-term impact all of you have experienced relating to gunning for promotions vs not doing so ov
reddit:ExperiencedDevs<!-- SC_OFF --><div class="md"><p>Why is it we only get GPU limited performance vs GPU data for new game releases and not CPU limited performance vs CPU? Is it simply because GPUs are easy to swap and PCIe is backwards compatible? Whereas a different CPU requires essentially enti
reddit:hardware<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1up1jq3/you_can_now_use_your_sony_headphones_as_a_free/"> <img alt="You can now use your Sony headphones as a free real-time head tracker for race and flight simulators on PC, several hundred games already suppo
reddit:hardware<!-- SC_OFF --><div class="md"><p>Here is how I do it:</p> <ol> <li>Find a some small scope that is relevant to my work. A recent example is SSO workflow in an open source project we use at work. I make sure to always pick a "leaf" for this task and not a trunk like ind
reddit:ExperiencedDevs  submitted by   <a href="https://www.reddit.com/user/sr_local"> /u/sr_local </a> <br /> <span><a href="https://www.tweaktown.com/news/112489/lenovo-starts-shipping-laptops-with-chinese-made-ymtc-ssds-due-to-skyrocketing-nand-flash-prices/index.html">[link]</a></span> 
reddit:hardware<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1uosum1/samsung_foundry_returns_to_profit_in_june_for_the/"> <img alt="Samsung Foundry Returns to Profit in June for the "First Time in Three Years"" src="https://external-preview.redd.it/LfPeUVKISW_yC
reddit:hardware<!-- SC_OFF --><div class="md"><p>Hi everyone,</p> <p>I'm a 2026 ECE graduate and have a VLSI startup offer (less than 10 LPA first year, around 10-13 LPA afterwards).</p> <p>I'm interested in Design Verification and eventually want to work at top semiconductor companies. I'm con
reddit:semiconductors<!-- SC_OFF --><div class="md"><p>In discussions about China’s semiconductor industry, the narrative is often reduced to two extremes: either “completely behind” or “about to surpass.” But if we break the industry down properly, the reality is far more nuanced. China has already
reddit:semiconductors<!-- SC_OFF --><div class="md"><p>A thread for Developers and IT folks with less experience to ask more experienced souls questions about the industry.</p> <p>​</p> <p>Please keep top level comments limited to Inexperienced Devs. Most rules do not apply, but keep it civil.
reddit:ExperiencedDevs<table> <tr><td> <a href="https://www.reddit.com/r/hardware/comments/1uogpo4/us_navy_is_flighttesting_3d_printed_fighter_jet/"> <img alt="US Navy is flight-testing 3D printed fighter jet parts that cut repair times in half — forward-deployed 3D printers generate composite parts,
reddit:hardware- 2026-07-05Floating gates are the only analog compute element you can fab today with zero extra mask steps
<!-- SC_OFF --><div class="md"><p>Some might say I'm obsessed with this whole analog computing thing. But its only because I find it so interesting! For some quick background I wrote a previous introductory explainer about how computing some things in the analog domain can be far
reddit:semiconductors - 2026-07-05Graviton G5 STREAM memory bandwidth
<!-- SC_OFF --><div class="md"><p>STREAM memory bandwidth results on an c9g.48xlarge instance.</p> <p>For reference my own HPC code (memory bandwidth limited with most time spent doing parallel sparse linear algebra) ran about 40% faster compared to the last generation Graviton (
reddit:hardware <!-- SC_OFF --><div class="md"><p>I have an old Xeon rig with 512Gb of 4-channel DDR4 2133 memory and E5-2699v4 processor. For GPU I have GTX 1060 with 6Gb of VRAM, so I use CPU only mode. I can run GLM 5.2 with 40B active parameters in Q4_K_XL at 1.8 t/s, but as you can understa
reddit:LocalLLaMA- 2026-07-05A planetary test for local models
<!-- SC_OFF --><div class="md"><p>Here is a fun test prompt:</p> <p>Imagine a date in the next 1000 years where the Sun, along with its gravity, suddenly disappeared. When that happens, all planets in our solar system would stop orbiting and carry in a straight line. Is there a d
reddit:LocalLLaMA - 2026-07-055060 worth it?
<!-- SC_OFF --><div class="md"><p>I had built a pc in 2023 for pretty cheap, it has a RTX 4090 and intel i9 13900K. I’m thinking of adding a 5060Ti for additional LLM workloads. I primarily use my GPU for training SLMs and llama.cpp inference.<br /> Since the memory bandwidth of
reddit:LocalLLaMA - 2026-07-05I benchmarked 13 models at 65K-128K context to find out what actually matters for agentic workloads
<!-- SC_OFF --><div class="md"><h1>I benchmarked 13 models at 65K-128K context to find out what actually matters for agentic workloads — prefill dominates everything, and KV head count beats parameter count</h1> <p>I've been running local LLMs for agentic workflows (tool use, cod
reddit:LocalLLaMA - 2026-07-05Concurrency plus nvfp4 on Blackwell
<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1unqkjy/concurrency_plus_nvfp4_on_blackwell/"> <img alt="Concurrency plus nvfp4 on Blackwell" src="https://preview.redd.it/phylna3ajbbh1.png?width=140&height=127&auto=webp&s=9dabf768fd1b49bfd934f3c
reddit:LocalLLaMA <!-- SC_OFF --><div class="md"><p>Not really a tutorial, but more of sharing my attempts at getting higher contexts on Q8 of Qwen3.6-27 with 32GB VRAM.</p> <p><strong>Disclaimer</strong>: Not in-depth research. Crowd wisdom suggests that Qwen is more tolerant of model quantizatio
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1unobl4/using_applications_to_make_a_smaller_model_more/"> <img alt="Using "applications" to make a smaller model more effective at bigger tasks." src="https://external-preview.redd.it/a2trYXYycGQxYm
reddit:LocalLLaMA- 2026-07-04Appreciation post!
<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1unn4j3/appreciation_post/"> <img alt="Appreciation post!" src="https://preview.redd.it/xokoffoerabh1.png?width=640&crop=smart&auto=webp&s=f2b1fafaf2adf5fb4935393cf2497f090bca41da" title="Appreciat
reddit:LocalLLaMA <!-- SC_OFF --><div class="md"><p>We moved our agent fleet's working memory off Markdown and onto TOON (Token-Oriented Object Notation) in December 2025 and just wrote up what 14 harnesses taught us.</p> <p>The honest numbers (tiktoken o200k, 100 uniform CRM records):</p> <p>- TO
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1unif51/possible_evidence_of_literal_prompt_injection_by/"> <img alt="possible evidence of literal prompt injection by anthropic" src="https://external-preview.redd.it/72by0vodw29h1.png?width=640&crop=smar
reddit:LocalLLaMA<!-- SC_OFF --><div class="md"><p>I'm using the model with opencode and the issue is it's looping hard when reasoning. It's not a deranged babbling though, the reasoning is legit large spans of text, it just can't get outside of the loop and make a decision. So I babysit it, stop
reddit:LocalLLaMA<!-- SC_OFF --><div class="md"><p>Is this a doable build? I know that it is possible to overflow models from the memory, so I'm wondering if a q4 or q6 quantization could work with this set up. Anyone running anything similar, with memory offload of models?</p> </div><!-- SC_ON -
reddit:LocalLLaMA- 2026-07-04Gemma 4 12B - MLX Kernel
<!-- SC_OFF --><div class="md"><p>I've mentioned this kernel project I was working on in a few posts and figured I would just open the project code for anyone curious: <a href="https://github.com/jscott3201/Helios">MLX Gemma 12B</a></p> <p>The main constraints for this on my end
reddit:LocalLLaMA <!-- SC_OFF --><div class="md"><p>Check it out: <a href="https://github.com/fairydreaming/llama.cpp/tree/dsv4">https://github.com/fairydreaming/llama.cpp/tree/dsv4</a></p> <p>They are PRs <a href="https://github.com/ggml-org/llama.cpp/pull/25247">#25247</a>, <a href="https://gith
reddit:LocalLLaMA<!-- SC_OFF --><div class="md"><p>Found cheap one RTX5000 16GB is it worth to add to RTX3060? And is it difficult to make them work together in LM Studio? Do I need to use studio drivers and is RTX5000 still supported with driver updates? Or should I just go for 5060Ti instead? T
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1unbn4y/comparing_local_inference_speeds_across_a_few/"> <img alt="Comparing local inference speeds across a few real setups people are running (3090 vs 5090 vs dual 6000)" src="https://preview.redd.it/8aep4zi
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1unbm45/ran_a_classicmedival_europe_fantasy_rpagentic/"> <img alt="Ran a classic(medival europe) fantasy RP/agentic benchmark across 8 local models Qwen3.6-27B held up better than its size suggests" src="https
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1unbjr2/using_local_models_with_hermes_vs_claude_code/"> <img alt="Using local models with Hermes vs Claude code" src="https://preview.redd.it/ii7l8k4wa8bh1.png?width=640&crop=smart&auto=webp&s=350
reddit:LocalLLaMA<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1unbi4a/qwen36_27b_on_a_5090_64k_sample_toks_distribution/"> <img alt="Qwen3.6 27B on a 5090, 6.4k sample tok/s distribution after tuning MTP/cache settings" src="https://preview.redd.it/l2k7gu5cb8bh1.jpeg?wid
reddit:LocalLLaMA- 2026-07-04DGX Spark and Overtemps
<!-- SC_OFF --><div class="md"><p>For anyone who has a DGX-Spark and is having problems during these very hot summer months, you can underclock with:</p> <p>sudo nvidia-smi -lgc 0,900</p> <p>My temps dropped from 85C to 60C and this fixed my problem of overtemp lockups.</p> <p>Ed
reddit:LocalLLaMA - 2026-07-04Feedback needed
<!-- SC_OFF --><div class="md"><p>Dear LocalLLaMA,</p> <p>I’m looking for feedback on my project because I noticed it is mainly catered towards technical people.</p> <p><a href="https://github.com/heterodoxin/apostate">Apostate</a> (Apostate is an abliteration engine built from s
reddit:LocalLLaMA <!-- SC_OFF --><div class="md"><p>A lot of people seem to be confused or mystified about this so figured I'd spell it out.</p> <p>I played around with RYS and realized that it broke Gemma 4 models. Turns out there's a `layer_scalar` value that is applied at each layer. If you don
reddit:LocalLLaMA<!-- SC_OFF --><div class="md"><p>Here's the setup I decided on for embedding gemma4-12b into a Tauri2 desktop app:</p> <ul> <li>Native Rust FFI into llama.cpp via <code>llama-cpp-2</code> (Metal enabled)</li> <li>Model:<code>gemma-4-12b-it-Q5_K_S</code> quantized by Unsloth, Q5_
reddit:LocalLLaMA- 2026-07-04First time I have seen this: my model seemed aware of its context usage ask me for compaction!
<!-- SC_OFF --><div class="md"><p>I was in a middle of a Claude Code session with GLM 5.2. Context usage 537k/1M. After finishing a task, GLM asked me this:</p> <blockquote> <p><strong>Context</strong> <strong>note:</strong> this session has run long and context is getting heavy.
reddit:LocalLLaMA <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1un9955/paper_gear_guided_endtoend_autoregression_for/"> <img alt="[Paper] GEAR: Guided End-to-End AutoRegression for Image Synthesis" src="https://preview.redd.it/vo0a7q0ut7bh1.png?width=640&crop=smart&am
reddit:LocalLLaMA- 2026-07-04Using structured insurance/captives to solve public pushback & zoning delays? (Idea discussion)
<!-- SC_OFF --><div class="md"><p>I’m looking for some blunt feedback on a strategy to handle the increasing community/municipal pushback on new builds—specifically regarding power grid strain, water use, and the "what's in it for the community" argument.<br /> Right no
reddit:datacenter - 2026-07-04Questions about AWS data center roles.
<!-- SC_OFF --><div class="md"><p>Hello everyone, I'm currently working at a data center on the operations side of things and am looking to move over to the compute side, something I can't do with my current employer. </p> <p>I'm wondering what are the differences between an Infr
reddit:datacenter