Skip to content
  • Facebook
  • Twitter
  • Instagram
  • Email
NEWSX 360 : The Global News Platform

NEWSX 360 : The Global News Platform

News from across the Globe

  • Home
  • About Us
  • Terms & Conditions
  • Contact Us
  • Editor’s Note
  • Privacy Policy
  • Toggle search form
  • Webapps Software Solutions is Indeed Making Technical Visions a Reality News
  • Singer Yashita Sharma Hits The Right Note For The Wedding Season With Her New Single “Bidaai” News
  • Mission Unicorn’s ‘Leaders of Tomorrow’: Empowering Entrepreneurs to Secure Funding and Thrive News
  • The Shining Media is expanding its network by launching “Business Headline” for Targeted Audiences News
  • Nagaland women nominated at National road safety council News
  • Pathumwadee Manu Engineering Opportunity Across Industries Borders and Limits India
  • Upcoming Residential Project “ARIA” BY NAMISHREE, Hyderabad News
  • Poor or very poor air quality in the city News

Guardrails for the Frontier: How AI Safety is Actually Being Built

Posted on May 2, 2026 By NewsX 360 No Comments on Guardrails for the Frontier: How AI Safety is Actually Being Built

In January 2026, Dario Amodei wrote a 20,000 word essay that made waves across the internet. The CEO of Anthropic, one of the tech giants and leaders in AI, has been openly talking about safety issues and risks associated with this emerging technology. In his essay, The Adolescence of Technology, Dario wrote in depth about various anticipated risks and the need for private organizations and governments to work together in forming policies, laws, and systems to mitigate these risks. He also took a dialectical approach, arguing that the positive impact of AI could far outweigh the risks associated with it.

In the past two years, with the speed of AI development, we have seen governments and private organizations take action through reforms and internal processes to control both current and anticipated threats. The most common approach is to identify, evaluate, and mitigate these risks.

How we categorize and assess AI risks

The most comprehensible threats associated with AI today are Biological/Chemical, Cybersecurity, Manipulation, and Model Autonomy.

● Biological and Chemical: This includes nuclear and radiological hazards where AI enables a non-expert to develop known or unknown biological threats.
● Cybersecurity: Offense can be done by enhancing the effectiveness of human attackers or through autonomous cyberattacks executed by AI systems from start to finish.
● Manipulation: This refers to the ability of AI systems to influence human beliefs, potentially influencing individuals to act against their own interests which could implicate democratic processes, social stability, and information integrity.
● Model Autonomy: The ability of AI systems to act independently, where they could replicate, survive, or conduct research to improve their own capabilities.

Then there are unknown risks that we may not comprehend or anticipate right now but may arise in the future. These threats are achievable when adversaries (individuals or well-resourced organizations) get unauthorized access to model weights or misuse the technology to exploit vulnerabilities.

Currently, most of the leading developers are setting multiple security thresholds depending on the model, an alarm is raised if a threshold is exceeded. Anthropic introduced ASL (AI Safety Levels), where each level requires specific safeguards. Google DeepMind uses Critical Capability Levels (CCLs) to represent points where AI may pose heightened risks. OpenAI tracks risks through defined categories with a gradation scale ranging from low to critical.

How we move from risk to regulation

The EU was the first to introduce the “EU AI Act”, governing AI development based on risk categories: unacceptable, high, limited, and minimal. In the US, the New York legislation introduced the Responsible AI Safety and Education Act (RAISE Act) to govern “frontier models.” California also introduced the Transparency in Frontier Artificial Intelligence Act (TFAIA) in Sep’25, targeting large developers for accountability. Both provide whistleblower protections and include significant financial penalties for non-submission of reports or disclosure of risks. These frameworks focuses majorly on frontier models that are more prone to systemic risks.

In the private sector, tech giants have taken regulation into their own hands. Google DeepMind has the Frontier Safety Framework, Anthropic regularly updates its Responsible Scaling Policy (RSP), and OpenAI has its Preparedness Framework. While distinct, they all share common steps: identifying, evaluating, mitigating, and governing risks. Currently, companies use methods like red teaming to stress-test models at different levels of development and deployment.

Securing access to model weights is one of the most critical safety norms among AI developers. Other key policies include the reporting of risks, rigorous third-party audits, and the tracking of incidents and mitigation for future reference. These frameworks ensure that crossing risk thresholds triggers immediate, non-discretionary actions such as halting deployment or hardening physical security. However, these regulations need to move from unilateral, company-led measures toward a coordinated multilateral ecosystem, where transparency and shared information flow among all stakeholders ensures that AI progress does not outpace our collective ability to control it.

About Author SHRUTI RAJVANSHI, Associate Director | Market Xcel

Shruti Rajvanshi is an Associate Director at Market Xcel and a postgraduate student at the Georgia Institute of Technology (MS, Human-Computer

Interaction). She holds a Bachelor’s degree in Computer Science from the University of Delhi and an MBA in Marketing and Finance. With over a decade of experience across analytics and business strategy, she has independently built AI-driven applications end-to-end using large language models, and has worked across emerging technology ecosystems including cryptocurrencies and NFTs. Her work reflects a strong, hands-on engagement with how technology is being built and deployed in real-world systems.

The post Guardrails for the Frontier: How AI Safety is Actually Being Built appeared first on The Blunt Times.

Related

India, News

Post navigation

Previous Post: ‘Shahid Smriti Van’ Validated as Key Pollution Mitigator in National Study
Next Post: WATCH: Punjab Kings Fielder Shashank Singh Drops Easy Catch During Practice Session; VIDEO Viral

Related Posts

  • Breaking Digital Frontiers: Internet Commerce Summit 2023 Set to Unveil the Future of E-Commerce News
  • निफ्टी-50 में जोमैटो-जियो फाइनेंशियल की एंट्री होगी:पेट्रोल-डीजल के दाम में आज कोई बदलाव नहीं, PM के प्रिंसिपल सेक्रेटरी बने RBI के पूर्व गवर्नर शक्तिकांत दास India
  • Fintifi Launches as a New Digital Platform Simplifying Access to Loans and Credit Cards with Advanced Technology News
  • Surat Celebrity Box Cricket League Season 4: A Mega Platform Uniting Influencers and Brands Nationwide for Collaborative Success News
  • BollyBeats® – The Fastest Growing Fitness Cardio Dance Workouts And Exercise Program In Asia News
  • The New Oaplus Series of Designer Lever Handles by Hafele News

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

  • Art of Living
  • Arts
  • Auto & Transportation
  • Automobile
  • Aviation
  • Banking
  • BCA ELECTIONS 2026
  • Bollywood
  • Brands
  • Business
  • Business Technology
  • Commodities
  • Economy
  • Education
  • Energy
  • Entertain­ment & Media
  • Entertainment
  • Entrepreneurs
  • Environment
  • Financial Services & Investing
  • Fitness
  • Gadgets
  • Health
  • Housing & Infrastructure
  • India
  • Information Technology
  • International
  • International Education
  • Investment
  • Lifestyle
  • Narendra Modi
  • News
  • People & Culture
  • Pharma
  • Policy & Public Interest
  • Politics
  • Sports
  • Stock Market
  • Tamilnadu
  • Technology
  • Telecom
  • The Multitaskers: A Series On Entrepreneurship with difference.
  • Travel
  • Uncategorized
  • UNICEF
  • Wellness
  • World News

Recent Posts

  • RBI Keeps Repo Rate Unchanged at 6.5%; Focus Shifts to Inflation Trajectory
  • Gold Prices Cross ₹78,000 per 10 Grams; Silver at 8-Month High
  • Sensex Hits Record High of 82,500 as FIIs Pour Record ₹18,000 Cr in a Single Day
  • Tamil Nadu CM Vijay Appoints Personal Astrologer Radhan Pandit Vetrivel as Officer on Special Duty
  • UP Census Enumeration Begins May 7 In Two Phases; Final Population Data To Be Based On March 1, 2027 Midnight

Recent Comments

No comments to show.

Archives

  • June 2026
  • May 2026
  • April 2026
  • March 2026
  • February 2026
  • January 2026
  • December 2025
  • November 2025
  • October 2025
  • September 2025
  • August 2025
  • July 2025
  • June 2025
  • May 2025
  • April 2025
  • March 2025
  • February 2025
  • January 2025
  • December 2024
  • November 2024
  • October 2024
  • September 2024
  • August 2024
  • July 2024
  • June 2024
  • May 2024
  • April 2024
  • March 2024
  • February 2024
  • January 2024
  • December 2023
  • November 2023
  • October 2023
  • September 2023
  • August 2023
  • July 2023
  • June 2023
  • May 2023
  • April 2023
  • March 2023
  • February 2023
  • January 2023
  • December 2022
  • November 2022
  • October 2022
  • September 2022
  • August 2022
  • July 2022
  • June 2022
  • May 2022
  • April 2022
  • October 2021
  • September 2021

  • “IVF is Not the Last Resort – Boost Your Fertility Naturally,” Says Holistic Wellness Expert Brands
  • Orient Green Power Reports Robust 446% YoY Jump in Q1 FY26 Net Profit Brands
  • Yokogawa Launches OpreX Plant Stewardship Brands
  • How QLead.ai Boosted Admissions for an EdTech Leader with Voice-Verified Leads Brands
  • Dhillon Freight Carrier Limited to Launch INR 10.08 Crore IPO on BSE SME Brands
  • Palladian Partners Powers Record-Breaking Festive Launch — Pearl Icon by Chandiwala Group Sold Out in Just 2 Hours Brands
  • Janaawar: The Beast Within Raises The Bar for Indian Web Series with Suspense on ZEE5 Brands
  • Sattvik Certifications Launches Mobile App in Southeast Asia to Empower Ethical Consumption Brands

Powered by PressBook News WordPress theme