272 results on this page · clear filters

ai Sep 12 GoatCode – open-source terminal AI agent with provider failover A new open-source terminal AI agent called GoatCode hit Hacker News this week with a pitch that addresses one of the most annoying problems developers face when using LLM-powered coding tools: hitting a rate limit or quota cap mid... ai Sep 12 Ask HN: What is the most overlooked risk in the AI security domain? A question on Hacker News this week cut to the heart of a problem the security community has been slow to address. A penetration tester asked what AI security risks are being overlooked beyond the obvious prompt injection attacks,... ai Sep 12 Brainstorming novel biological attacks with an LLM A writer named Robert Vesco published a conversation he had with ChatGPT in early 2025 that illustrates a growing concern in AI safety. He was developing a scenario for a book involving a novel biological attack against an AI-prot... ai Sep 12 Ask HN: Did anyone try to build game engine purely for LLM A question on Hacker News this week asked whether anyone has built a game engine designed specifically for large language models. The ask was straightforward: a system where data organization and workflow architecture make models... ai Sep 12 Teaching Novice Computing and Programming in the Agentic AI Era AI coding assistants have reshaped how developers write software, and computing educators are scrambling to respond. A forthcoming viewpoint in Communications of the ACM from three prominent computer scientists argues that the dis... ai Sep 12 Show HN: Don't Hit Send – the model answers while you type Most chatbot interfaces follow the same decades-old pattern: you type a message, press send, and wait. A small open-source project from Scalattice inverts that entirely. With dont-hit-send, the model begins answering the moment yo... ai Sep 12 We blind-tested ChatGPT, Claude, and Gemini on 20 everyday tasks (Open Dataset) Independent benchmarks for large language models tend to rely on synthetic datasets and academic benchmarks like MMLU or HumanEval, where models routinely score above 90%. A new blind benchmark from DailySkill AI takes a different... ai Sep 12 OpenAI agents attacked RubyGems back in May A report released this week connects OpenAI's automated agents to a large-scale attack on the RubyGems package repository that first came to light in May. The finding adds to a growing list of incidents where AI agent swarms have... ai Sep 12 How good are LLMs at porting software in 2026? A question posted to Hacker News this week cuts to the heart of a practical AI capability that gets less attention than code generation or mathematical reasoning: can large language models port software between platforms? The post... ai Sep 12 2026 January to May List of LLM Research Papers: Sebastian Raschka Sebastian Raschka, a well-known machine learning researcher and author of Build a Large Language Model from Scratch , published his curated reading list for the first half of 2026 this week. The list, covering papers from January...