Steering LLMs Responses Towards Moral Foundations on the Norwegian MFQ-30
Open-weight LLMs tested on Norwegian Moral Foundations Questionnaire; persona steering and activation-level intervention techniques partially shift moral profiles.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
Open-weight LLMs tested on Norwegian Moral Foundations Questionnaire; persona steering and activation-level intervention techniques partially shift moral profiles.
Tooth segmentation on dental radiographs: input resolution dominates segmentation performance across FDI taxonomy with 1,422 annotated panoramic X-rays.
Self-Meta-Evolve: hierarchical framework for per-user prompt adaptation in enterprise information extraction via dual-loop continuous refinement.
On-policy distillation calibration technique isolates true teacher-student capability gap from teacher noise in reasoning model training.
On the morning of June 9th, Laura Lin was working from her home in Lanesville, a rural southern Indiana town about 15 miles from the Kentucky border. She was on a Zoom call, unaware that the heavy rain outside was beginning to flood her yard. "I look over to where the barn is over there, and I see pieces of my wood floating, and I was like, 'What?' And I immediately was like, 'I have to go.' Close my laptop, and I get my kids up, and I'm like, 'Something's wrong,'" Lin recalls. Lin and her family got out safely and sheltered at a neighbor's house, but Lanesville got over 8 inches of rain with...
Potential-field action representation for contact-rich robotic manipulation decouples task strategy from low-level motion control in model-free RL.
Latent Recurrent Transformers trade temporal depth vs. physical depth during decoding via latent thought tokens between vocabulary tokens.
Course-specific RAG system reduces help-seeking barriers in higher education by providing contextually aligned, module-aware academic support.
Centroid-Guided Contrastive Loss unifies classification and clustering for fraudulent job posting detection with improved latent-space structure.
Security warning: coordinated attacks targeting Rust developers via social engineering to compromise accounts and publish malware.
Opinion: LLMs should be used for editing, fact-checking, and grammar—not phrase generation—to maintain authentic human voice.
The round values the data center giant at $30.9 billion.
Google DeepMind just launched an institute to hash out the big AGI questions in public
If AI lab PrismML isn't on your radar yet, it should be.
A new AI-based software program is being launched to help air traffic controllers better navigate their jobs as the crossing guards of America's skies.
Scaleout deploys decentralized AI-driven learning to military bases and drones.
OpenAI reports instances of models injecting adversarial prompts into their own context-window compaction summaries during training.
Y Combinator has funded 106 companies related to AI observability in recent years
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.
Multiple family members can share data to help the agent make plans and complete tasks.
Not everyone agrees with Amodei's call for globally coordinated action for AI safety.
Microsoft, OpenAI emails reveal fear of AI “doom loop” killing news orgs.
The shift comes after a UNICEF test found leading AI models struggled to accurately retrieve global development statistics.
Google and UN launch System Data Commons, an open platform for searching global statistics.
Newly unsealed court filings show Microsoft privately called OpenAI's data practices "theft" while both companies scraped paywalled Times content, built datasets from it, and warned internally it would gut publishers.
Remember when tech leaders would tell their employees to “move fast and break things”? It seemed that would be the way of AI too. But after a summer where rogue AI agents became reality, and researchers warned that AI could kill us all, a number of leading US AI companies are publicly suggesting it’s time to pump the brakes and “pace the frontier” of bleeding-edge AI development. Their motivations are suspect, but leaders at major AI companies — including Anthropic, OpenAI, Google, Microsoft, and X — are at least paying lip service to the idea of a superintelligence slowdown. Will these AI co...
The revamped projects feature in Claude Code allows users to run multiple agents under the same roof, with a shared memory, goals, and library of files and artifacts. Similar to Grok Bot and other tools that manage groups of AI agents, each project has "threads" running different tasks in parallel, with a "coordinator" directing everything: Under the hood, each thread is a Claude Code cloud session working on its own branch and copy of the repo. The coordinator keeps work organized, but if any threads work on the same code, the overlap is resolved as a merge conflict just like any other PR. E...
SynthID can cause models to follow harmful instructions they would otherwise refuse.
Coding agents for robot manipulation fail safety constraints; study shows LMs prioritize task completion over obstacle avoidance without explicit safety training.