AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

AI safety

Explore collected AI evidence about AI safety, with dates and links to original sources.

Showing 20 of 632 matching collected records. Text matches can include mentions by other organizations.

  1. Sep 21, 2026 · UTC · NVIDIA Newsroom

    Why Deploying Physical AI at Scale Demands Safety at Every Layer

    Physical AI is moving rapidly from research to large-scale deployment. By 2035, ABI Research projects an installed base of 49 million level 3-5 autonomous vehicles (AVs), while Omdia estimates that roughly 60 million industrial robots will be deployed between 2026 and 2035. As these machines enter roads, factories, warehouses and other environments shared with people, […]

  2. Sep 21, 2026 · UTC · OpenAI News

    Building standards for the next phase of AI

    OpenAI outlines a path to shared global AI standards, calling for coordinated evaluation, reporting, and governance to improve safety.

  3. Sep 20, 2026 · UTC · The Verge AI

    Humans, not rogue AI, are still the biggest cybersecurity risk to energy systems

    Before recent high-profile hacks raised the specter of AI possibly "killing all humans," our energy systems were already disturbingly vulnerable to cyberattack - and the risk is growing. "We were always prey. We were just kind of surviving at the appetite of our predators," Joshua Corman, executive in residence for public safety and resilience at […]

  4. Sep 19, 2026 · UTC · TechCrunch AI

    AI safety conversations have gotten unbelievable

    This week two conversations about AI safety went viral that demonstrate just how hard it is to discern AI fact from fiction.

  5. Sep 19, 2026 · UTC · OpenAlex research metadata

    Using Multimodal Large Language Models for False Alarm Reduction in Image-Based Fire Detection

    Fire Technology · Fire Detection and Safety Systems · University of Science and Technology of China

  6. Sep 19, 2026 · UTC · OpenAlex research metadata

    Artificial intelligence and predictive analytics for H₂S risk management in high-hazard oil and gas facilities

    Актуальные исследования · Risk and Safety Analysis

  7. Sep 19, 2026 · UTC · OpenAlex research metadata

    Case-knowledge-guided construction-phase safety appraisal report generation for hydraulic and hydropower projects using large language models

    Engineering Applications of Artificial Intelligence · Occupational Health and Safety Research

  8. Sep 18, 2026 · UTC · TechCrunch AI

    Dario Amodei and other AI leaders want to ‘Pace the Frontier’ but…how?

    A week after an Anthropic researcher’s doomsday warning rattled the AI world, the company’s CEO Dario Amodei has outlined his plan to “pace the frontier” of AI development. The proposal leans on independent safety evaluators and coordination between AI labs in democratic countries, and it’s already picked up some industry support, along with some pointed pushback from Nvidia’s Jensen Huang. Watch […]

  9. Sep 18, 2026 · UTC · TechCrunch AI

    Automattic’s 33-Hour Coup, and can AI labs police themselves?

    A week after an Anthropic researcher’s doomsday warning rattled the AI world, the company’s CEO Dario Amodei has outlined his plan to “pace the frontier” of AI development. The proposal leans on independent safety evaluators and coordination between AI labs in democratic countries, and it’s already picked up some industry support, along with some pointed pushback from Nvidia’s Jensen Huang. On […]

  10. Sep 18, 2026 · UTC · OpenAI News

    Introducing the Australian Youth Safety Blueprint

    OpenAI introduces the Australian Youth Safety Blueprint, a six-pillar roadmap for safer AI experiences that protect and empower young people.

  11. Sep 17, 2026 · UTC · arXiv · Artificial Intelligence

    Coding Agents with an Obstacle-Aware Harness for Safe Robot Manipulation

    Coding agents have emerged as a promising paradigm for robot manipulation: a language model writes the robot controller as a program, and agents built in this way now operate robots without robot-specific training.Whether this paradigm is also safe, however, has not been asked. We evaluate coding agent under a safety constraint, where each task pairs a manipulation goal with an obstacle the robot must not touch. The agent pursues the goal but collides with the obstacle in most cases, treating task completion as its sole objective while neglecting safety. The agent reasons about the obstacle in

  12. Sep 17, 2026 · UTC · arXiv · AI, language, vision and robotics

    Harm Laundering in GPT Models: Evidence That Gender Discrimination Is Transformed Rather Than Reduced Across Safety-Trained Generations

    Safety evaluations for large language models rely on surface-form classifiers that report declining harm scores across model generations. We provide evidence that this methodology is systematically incomplete: explicit discriminatory content is transformed rather than removed. We call this \emph{harm laundering}. Analysing 450,000 gender-directed completions across 15 models spanning GPT-2 through to GPT-5 (OpenAI GPT lineage; three demographic conditions), we show that sexual violence clusters prevalent in GPT-2 women-directed output disappear by GPT-4, while men-directed completions gain pos

  13. Sep 17, 2026 · UTC · arXiv · AI, language, vision and robotics

    OPTED: On-Policy Fine-Tuning for End-to-End Driving using a Render-Free Teacher

    As scaling pre-training data alone yields diminishing returns, post-training is becoming increasingly important across physical AI domains such as autonomous driving. End-to-end driving policies are pre-trained in open loop with behavior cloning on human demonstrations. However, compounding errors during closed-loop deployment can take the vehicle outside the training data distribution, increasing the risk of safety-critical incidents. Closed-loop post-training can mitigate this risk but requires costly simulation for sensor-based policies. We propose OPTED (on-policy fine-tuning for end-to-en

  14. Sep 17, 2026 · UTC · arXiv · AI, language, vision and robotics

    Custom PX4 firmware for autonomous hybrid aerial-marine missions

    Mapping and monitoring aquatic environments can benefit from hybrid aerial-amphibious drones able to combine flight and water-surface navigation within the same mission. This paper presents a PX4 firmware extension for such platforms, introducing manual and autonomous marine navigation modes integrated with the standard PX4 mission pipeline and QGroundControl interface. The proposed framework preserves existing flight functionalities and safety mechanisms while enabling unified planning and execution of hybrid aerial-marine missions with differentiated aerial and marine waypoints. Simulated ca

  15. Sep 17, 2026 · UTC · arXiv · AI, language, vision and robotics

    SAFARI: An Industrial Benchmark for LLM-Assisted Hazard Analysis and Risk Assessment

    Large language models (LLMs) are increasingly considered for safety-critical engineering, yet their reliability in regulated functional-safety workflows remains underexplored. We introduce SAFARI (Safety-Aware Functional Automotive Risk Inference), the first industrial benchmark for LLM-assisted automotive Hazard Analysis and Risk Assessment (HARA) under ISO 26262. It contains 3,000 de-identified industrial HARA cases and evaluates two coupled tasks: open-ended hazard analysis and standards-grounded risk assessment. To evaluate open-ended HARA artifacts, we propose the first reference-anchored

  16. Sep 17, 2026 · UTC · arXiv · AI, language, vision and robotics

    NS3Learn: Transferring 5G NR Mode-2 Reception Realism from ns-3 to the Veins/SUMO Stack for Connected-Vehicle Safety Assessment

    Connected-vehicle safety evaluations rely on coupled traffic and network simulations, but standard channel models ignore radio resource competition in 5G NR sidelink Mode-2, reporting unrealistically high message delivery in dense traffic. This study introduces resource-competition losses without requiring full protocol reimplementation. We labeled 10.5 million reception outcomes from ns-3 5G-LENA traces (calibrated on 3GPP scenarios and driven by SUMO trajectories) to fit NS3Learn - a closed-form model capturing half-duplex loss, scheduling collisions, receiver capture, and decoding. Evaluati

  17. Sep 17, 2026 · UTC · AWS Artificial Intelligence Blog

    Enhancing industrial safety AI with synthetic data on Amazon SageMaker AI

    Learn how to build a synthetic data augmentation pipeline on Amazon SageMaker AI and Amazon Rekognition that generates photo-realistic, auto-labeled training images for industrial safety AI. This approach improved person detection by up to 160% without manual annotation or hazardous data collection near heavy machinery.

  18. Sep 17, 2026 · UTC · arXiv · AI, language, vision and robotics

    Time-Efficient Iterative Learning Planning for Safety-Critical Dynamic Obstacle Avoidance

    Autonomous mobile robots require timeefficient planning and safety-critical dynamic obstacle avoidance under constrained onboard computation. While Iterative Learning Planning (ILP) offers lightweight and efficient traversal planning, it lacks explicit mechanisms for dynamic obstacle perception and avoidance. This article extends ILP to safety-critical navigation in dynamic environments by integrating an anticipatory risk-blended control barrier function (ARB-CBF). The extended ILP learns traversal-speed and steering-bias profiles via a fractionalpower update based on local obstacle risk, gene

  19. Sep 17, 2026 · UTC · arXiv · Artificial Intelligence

    Local Sparsity Enables Unsupervised LLM Safety Detection

    Deployment-time safety methods for large language models (LLMs) are predominantly supervised and assume access to unsafe training data. Nevertheless, new attacks and harm categories regularly arise, not captured by models trained in such a supervised fashion. An alternative approach is to view this problem through the lens of anomaly detection, namely, to rely solely on modeling safe data and flagging out-of-distribution inputs. However, LLM activations lie in a high-dimensional space, raising concerns about whether anomaly detection is statistically feasible. We show that, under the linear re

  20. Sep 17, 2026 · UTC · arXiv · AI, language, vision and robotics

    QoS-Aware Federated Learning for Multimodal In-Cabin Interaction in Smart Vehicles

    Modern smart vehicles leverage multimodal sensors, ranging from high-bandwidth vision systems to low-rate physiological monitors, to provide personalized in-cabin services. However, integrating high-fidelity multimodal fusion with collaborative training is often hindered by the heterogeneous and time-varying Quality of Service (QoS) constraints of vehicular networks. Standard Federated Learning (FL) approaches enforce rigid synchronous rounds that fail to account for these resource asymmetries, leading to safety-critical timing violations and energy exhaustion. In this paper, we propose FedQoS

Explore full timeline