| 1. | HackerRank open sourced its ATS. My resume scored 90/100. Oh wait 74. No – 88(danunparsed.com) |
| 1012 points by sambellll 54 days ago | 432 comments | permalink | |
tl;dr: HackerRank's open-sourced ATS (hiring-agent) produces wildly inconsistent resume scores—the same resume scored anywhere from 66 to 99 across 100 runs, meaning candidates can fail an 85-point cutoff 65% of the time purely by luck. The author traces this to LLM non-determinism on subjective judgments (projects vary hugely, while checklist-style skills stay consistent) and a two-line "experience" prompt with no rubric that awards 25/25 to everyone from interns to principal engineers. The takeaway: LLMs are fine for parsing resumes but shouldn't be making qualitative scoring calls that decide who gets filtered out. | |
HN Discussion:
| |
| 2. | GLM 5.2 beats Claude in our benchmarks(semgrep.dev) |
| 1097 points by jms703 54 days ago | 504 comments | permalink | |
tl;dr: Semgrep benchmarked open-weight and frontier models on detecting Insecure Direct Object Reference (IDOR) vulnerabilities, and found that Zhipu AI's GLM 5.2 scored 39% F1 with just a bare prompt, beating Claude Code (32%) at roughly $0.17 per vulnerability found and ~1/6 the cost. However, both trailed Semgrep's own purpose-built multimodal pipeline (53–61% F1), reinforcing that the harness around a model matters more than the model itself. Other open-weight models (MiniMax M3, Kimi K2.7) lagged significantly, so GLM 5.2 appears to be a standout rather than evidence of open weights broadly catching up. | |
HN Discussion:
| |
| 3. | Age verification is just a precursor to automated attribution of speech(nonogra.ph) |
| 954 points by arkhiver 54 days ago | 595 comments | permalink | |
tl;dr: Age verification laws, marketed as child protection measures, are actually identity attribution systems that link online accounts to real-world identities like SSNs and government IDs. This solves law enforcement's traditionally labor-intensive problem of identifying anonymous speakers, potentially enabling automated enforcement against inconvenient speech once adoption is widespread. The author urges readers to refuse verification, or if unavoidable, use third-party services paid in Monero. | |
HN Discussion:
| |
| 4. | Historical memory prices 1960-2026(dam.stanford.edu) |
| 397 points by vga1 54 days ago | 152 comments | permalink | |
tl;dr: An interactive dataset tracking historical memory prices ($/GB) from 1960 to 2026 across DRAM, NAND flash, and HBM, extending John C. McCallum's classic dataset with monthly updates from Keepa (Amazon retail) and quarterly HBM estimates from Epoch AI. It breaks out DRAM by generation (SDRAM through DDR5), HBM by generation (HBM2e through projected HBM4), and includes a modeled accelerator cost breakdown for Nvidia, AMD, Google TPU, and AWS Trainium. Raw CSV data is downloadable; caveats note that "cheapest retail" often reflects EOL clearance rather than leading-edge pricing. | |
HN Discussion:
| |
| 5. | 5k menus from the New York Public Library’s Buttolph Collection (1880-1920)(pudding.cool) |
| 408 points by xbryanx 54 days ago | 107 comments | permalink | |
tl;dr: Summary not available | |
HN Discussion:
| |
| 6. | I used Claude Code to get a second opinion on my MRI(antoine.fi) |
| 548 points by engmarketer 54 days ago | 681 comments | permalink | |
tl;dr: After an orthopedist diagnosed a Grade III partial-thickness subscapularis tendon tear and began aggressive treatment (including shockwave therapy and a homeopathic injection), the author ran their DICOM MRI files through Claude Code (Opus) for a second opinion, which instead found an intact tendon with only mild tendinosis. A follow-up arbitration run by Claude sided with the "no tear" reading at moderate-to-high confidence. The author acknowledges they can't fully trust either verdict but is now skeptical of the clinic's intervention-heavy plan. | |
HN Discussion:
| |
| 7. | Show HN: Zanagrams(zanagrams.com) |
| 372 points by pompomsheep 54 days ago | 100 comments | permalink | |
tl;dr: Summary not available | |
HN Discussion:
| |
| 8. | The KIDS Act would require age checks to get online(eff.org) |
| 624 points by bilsbie 54 days ago | 536 comments | permalink | |
tl;dr: Congress is fast-tracking the KIDS Act, a bundle that combines a revised KOSA with bills like SAFE BOTS and SCREEN, imposing liability whenever platforms "knew or should have known" a user is a minor—effectively forcing age verification (via ID or facial scans) for all users despite disclaimers to the contrary. The package also pressures platforms to moderate broad categories of lawful speech (addiction, gambling, fraud discussions) and contains encryption carve-outs with loopholes that could undermine private and ephemeral messaging. | |
HN Discussion:
| |
| 9. | Professor denounces mass AI fraud on an exam at Brown(english.elpais.com) |
| 527 points by geox 54 days ago | 689 comments | permalink | |
tl;dr: Brown University economics professor Roberto Serrano caught at least 50 students cheating on a take-home midterm using ChatGPT, after the class averaged 96/100 but dropped to 48/100 on an in-person final—with 22 of 27 no-shows having previously scored perfect 100s. Serrano criticized Brown's administration for responding with silence and is abandoning take-home exams, while urging broader debate on AI-enabled academic fraud. The incident reflects a wider trend: Princeton recently ended its 133-year-old unproctored honor code system in response to AI cheating. | |
HN Discussion:
| |
| 10. | Librepods: AirPods liberated(github.com) |
| 479 points by rbanffy 54 days ago | 175 comments | permalink | |
tl;dr: LibrePods reverse-engineers Apple's proprietary AirPods protocol to bring exclusive features—like noise control switching, ear detection, accurate battery status, conversational awareness, and head gestures—to Linux and Android. Spoofing a Vendor ID as Apple's unlocks additional capabilities such as accessibility settings and hearing aid customization, while features like Find My, spatial audio, heart rate monitoring, and high-quality two-way audio are planned but likely require root. The project is GPLv3-licensed and warns that librepods.org is an unofficial site falsely claiming affiliation. | |
HN Discussion:
| |
| 11. | A way to exclude sensitive files issue still open for OpenAI Codex(github.com) |
| 224 points by pikseladam 54 days ago | 142 comments | permalink | |
tl;dr: A GitHub issue requests that OpenAI Codex add a `.codexignore` mechanism (both repo-level and global) to explicitly prevent the agent from reading or transmitting sensitive files like `.env`, `.pem`, or SSH keys to the model. The requester notes this was previously raised in issue #205, which was closed in favor of a Rust implementation (codex-rs), but no equivalent feature appears to exist there as of August 2025. | |
HN Discussion:
| |
| 12. | The curious case of the disappearing Polish S (2015)(aresluna.org) |
| 252 points by colinprince 54 days ago | 106 comments | permalink | |
tl;dr: Medium users typing in Polish couldn't enter the letter Ś because the editor blocked Ctrl+S to prevent the browser's save dialog. The bug arose because Polish keyboards use Right Alt+S to type Ś, and Windows internally maps Right Alt to Ctrl+Alt—so the editor's Ctrl+S handler was swallowing the keystroke. The fix was a one-line change: only block Ctrl+S when Alt isn't also pressed. | |
HN Discussion:
| |
| 13. | EU to legislate about Chat Control behind closed doors(patrick-breyer.de) |
| 716 points by NeutralForest 54 days ago | 424 comments | permalink | |
tl;dr: Former MEP Patrick Breyer warns of a two-pronged EU push this weekend to revive "Chat Control" mass-scanning of private messages. EP President Metsola is reportedly trying to resurrect the expired Chat Control 1.0 regulation despite Parliament's March rejection, while Monday's trilogue on the permanent CSAR proposal could mandate warrantless scanning, "voluntary" detection as risk mitigation, and age verification that ends anonymous communication. Civil society has relaunched fightchatcontrol.eu to pressure lawmakers. | |
HN Discussion:
| |
| 14. | Marfa Public Radio Puts You to Sleep(marfapublicradio.org) |
| 414 points by reaperducer 55 days ago | 130 comments | permalink | |
tl;dr: Marfa Public Radio launched a sleep podcast called "Marfa Public Radio Puts You to Sleep" for its fall membership drive, in which staff read aloud the dull operational documents (FCC compliance, NPR ethics codes, etc.) that keep the 24/7 station running. The goal is both to lull listeners to sleep and to encourage donations at marfapublicradio.org/donate. | |
HN Discussion:
| |
| 15. | Michigan bill would bar employers from requiring after-hours coms with workers(cbsnews.com) |
| 279 points by cebert 54 days ago | 221 comments | permalink | |
tl;dr: Michigan Senate Bill 948, introduced by Sen. Erika Geiss, would prohibit employers from requiring workers to respond to emails, texts, or messages outside of their scheduled hours, with exceptions for contracted on-call pay, employee-set availability windows, and state/federal emergencies. Violations could be reported to the Department of Labor and Economic Opportunity, potentially resulting in fines or overtime pay. The bill has been referred to the Labor Committee. | |
HN Discussion:
| |
| 16. | OpenRA(openra.net) |
| 803 points by tosh 55 days ago | 163 comments | permalink | |
tl;dr: OpenRA's new playtest-20260222 introduces random map generators for Red Alert, Tiberian Dawn, and Dune 2000, usable in skirmish and multiplayer. Dune 2000 gets new visual effects, Starport bulk purchasing, and a community-led balance overhaul, while the standalone Tiberian Dawn HD mod is now feature-complete with toggleable remastered/classic assets. Other additions include map editor improvements, expansion-building bots, auto-save, new missions, and groundwork for localization. | |
HN Discussion:
| |
| 17. | Suspicious Discontinuities (2020)(danluu.com) |
| 276 points by tosh 55 days ago | 100 comments | permalink | |
tl;dr: Summary not available | |
HN Discussion:
| |
| 18. | Anonymous GitHub account mass-dropping undisclosed 0-days(github.com) |
| 941 points by binyu 55 days ago | 380 comments | permalink | |
tl;dr: Summary not available | |
HN Discussion:
| |
| 19. | DSpark: Speculative decoding accelerates LLM inference [pdf](github.com) |
| 791 points by aurenvale 55 days ago | 361 comments | permalink | |
tl;dr: Summary not available | |
HN Discussion:
| |
| 20. | AI learns the “dark art” of RFIC design(spectrum.ieee.org) |
| 270 points by Brajeshwar 58 days ago | 174 comments | permalink | |
tl;dr: Princeton researchers are using reinforcement learning, inverse design, and diffusion models to automate RFIC design—a notoriously artisanal field where chips for 5G, radar, and satellite comms have traditionally been hand-crafted over years. Their AI-generated power amplifiers, which often look like QR codes rather than symmetric layouts, have achieved record bandwidth and efficiency while cutting design time from months to minutes. The main bottleneck now is training data, most of which sits locked behind corporate NDAs, prompting calls for open chip-design datasets akin to ImageNet. | |
HN Discussion:
| |