Third-party cyber evaluations involving OpenAI models
OpenAI addresses third-party cybersecurity evaluation incidents and announces new safeguards for AI model testing protocols.
Every story tagged with this topic, ordered by date.
OpenAI addresses third-party cybersecurity evaluation incidents and announces new safeguards for AI model testing protocols.
Mariano-Florentino Cuéllar joins Anthropic as first Chief Global Affairs Officer, signaling focus on policy and governance.
Proposes security-oriented lifecycle model for LLM systems addressing provenance, signing, permissions, and decommissioning in critical infrastructure.
OpenAI responds to Apple lawsuit allegations, disputes claims about employees and shares internal communications.
Microsoft-led open letter signed by 235 AI companies including NVIDIA and OpenAI argues against US government restrictions on open-weight models on safety grounds.
Podcast discussion on open-weight model competitiveness, cybersecurity risks, and AI leadership policy letters signed by industry leaders.
OpenAI outlines safety, security, and transparency practices aligned with EU AI Act compliance and responsible governance.
Cohere signs EU AI Content Transparency Code, demonstrating early compliance with EU AI Act requirements.
Bruce Schneier argues writing assignments develop critical thinking skills that atrophy without practice, relevant to AI adoption decisions in education and work.
AISPA framework audits system prompts in LLM applications across 8 user-centric dimensions to address transparency gaps.
Framed behavioral experiment shows competitive AI race dynamics incentivize riskier development, validating speed-safety trade-off under falling-behind pressure.
Polistemics: theory-grounded benchmark for evaluating LLMs as political information mediators via epistemic modesty standard.
AgentToolMO proposes 3GPP NRM information model for cross-vendor AI agent tool trust management with graduated enforcement and cascade propagation bounds.
Policy analysis: general-purpose AI governance frameworks for public services risk failure due to GPAI properties (generality, accessibility, low cost) undermining safety preconditions.
Anthropic publishes official stance on open-weights model releases, addressing trade-offs between transparency, safety, and competitive positioning.
Survey of AI's role in innovation ecosystems covers digitization trends and macro economic integration.
Study finds Grok assigns 2-5x higher credibility to ethnonationalist pseudo-science than Claude, GPT, Gemini across four LLM families.
Theory paper argues human participation persists in automated systems for technical, complementarity, and normative reasons beyond current AI capability limits.
Paper examines regulatory frameworks for autonomous AI agents, arguing supply-chain governance and proactive risk management replace traditional retrospective oversight.
Anthropic outlines research priorities for its Economic Futures Research Fund, focusing on AI's labor and economic impacts.
Full-text AI detection on 14k Amazon self-published books (2023–2026) shows AI-heavy titles dominate catalog but underperform in sales.
232k dataset-model-app chains show license obligations stripped in AI supply chains; 'license laundering' quantified.
OpenAI announces Project Camellia infrastructure investment in Effingham County, Georgia with commitments to energy efficiency, job creation, and Codex access.
OpenAI partners with U.S. Department of Energy and national labs to apply frontier AI to scientific discovery and research acceleration.
Study models how imperfect LLM detectors distort user incentives and downstream metrics, showing counterintuitive effects of detection as behavioral intervention.
Anthropic donates additional $20M to Public First Action, totaling $40M commitment to AI policy advocacy.
Cohere partners with Government of Canada to develop sovereign AI capabilities for public-sector services.
David Vélez and Robin Vince appointed to OpenAI Foundation and OpenAI Group PBC boards.
Ben Thompson proposes US law to legalize model distillation and data collection as fair use, addressing licensing hypocrisy and competitiveness vs. Chinese models.
Import AI newsletter covers open vs closed model gaps, Kimi K3 release, and Demis Hassabis's AI policy proposals.
Stratechery argues U.S. frontier labs face minimal threat from Chinese models; policy should prioritize open-weight domestic alternatives instead.
Lightweight framework for auditable trustworthiness assessments in AI lifecycle governance with formal representation and monitoring.
Stratechery weekly digest covering mainframe obsolescence, OpenAI developments, and Netflix competitive position.
Methodology for harmonized AI safety thresholds across misuse, malfunction, and systemic risks to prevent race-to-the-bottom in standards.
Tutorial and survey on agentic AI for 5G/6G networks covering reasoning, planning, multi-agent coordination, and standardization.
Empirical evaluation finds LLM watermarking methods fail forensic standards required by EU AI Act and California SB 942.
Opinion piece arguing for independent AI certification mechanisms to address market failure in rewarding trustworthy development.
Formal semantics and reference implementation for ODRL policy evaluation in EU dataspaces and AI governance workflows.
PHP-AIO protocol formalizes hidden systemic risks (knowledge erosion, resilience, regulatory) in role-level automation decisions.
ArtSplit provotype explores how to quantify and attribute ownership between human and AI contributions in creative work through measurable contribution metrics.
Linus Torvalds states Linux will adopt AI tools; rejects anti-AI stance as maintainer policy.
Cohere partners with University of Toronto on multi-year AI adoption and responsibility initiative.
Formal differential-privacy framework for whistleblower protection against organizational retaliation via auditor-selection obfuscation.
OpenAI proposes 'reverse federalism' model where state-level AI regulations inform national safety framework.
User permission frameworks for AI agents address prompt injection and unauthorized actions; bridges product-level design with enforcement mechanisms.
Apple sues OpenAI over alleged trade secret theft; Stratechery frames it as reactive rather than substantive legal action.
Ben Bernanke joins Anthropic's Long-Term Benefit Trust, signaling governance structure for AI safety oversight.
Anthropic launches public Q&A initiative to address community concerns about AI safety and development practices.
Analysis extends agentic inequality framework: interaction-level context access (Dynamic Context) creates disparities independent of agent availability/quality.
Systematic review of agentic AI governance identifies governance priorities and mechanisms for autonomous planning/execution systems.