By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
AIModelKitAIModelKitAIModelKit
  • Home
  • News
    NewsShow More
    SpaceXAI’s Grok Tool Uploading Users’ Entire Codebase to Cloud Storage: What You Need to Know
    SpaceXAI’s Grok Tool Uploading Users’ Entire Codebase to Cloud Storage: What You Need to Know
    4 Min Read
    New York Leads the Way: First State to Enforce One-Year Moratorium on New AI Data Centers
    New York Leads the Way: First State to Enforce One-Year Moratorium on New AI Data Centers
    4 Min Read
    AI Replacing New York Nurses: Why Patients Should be Concerned About Quality of Care
    AI Replacing New York Nurses: Why Patients Should be Concerned About Quality of Care
    5 Min Read
    Navigating AI Agent Crawlers and Cloudflare’s New Rules: A Comprehensive Guide
    Navigating AI Agent Crawlers and Cloudflare’s New Rules: A Comprehensive Guide
    5 Min Read
    How Apple’s Self-Driving Car Program Paved the Way for Advanced AI Chip Technology
    How Apple’s Self-Driving Car Program Paved the Way for Advanced AI Chip Technology
    4 Min Read
  • Open-Source Models
    Open-Source ModelsShow More
    Unlocking Efficient Autoregressive Video Generation with SemanTok: Predictable Semantic Tokens by Stability AI
    Unlocking Efficient Autoregressive Video Generation with SemanTok: Predictable Semantic Tokens by Stability AI
    5 Min Read
    4Director: Mastering Video World Models with Rigid 3D Geometry | Stability AI Insights
    4Director: Mastering Video World Models with Rigid 3D Geometry | Stability AI Insights
    6 Min Read
    Leveraging Earth AI’s Geospatial Foundation Models to Enhance Global Public Health Initiatives
    Leveraging Earth AI’s Geospatial Foundation Models to Enhance Global Public Health Initiatives
    5 Min Read
    Enhancing AI Image Generation with Diffusion Controller: A Simplified Unified Approach
    Enhancing AI Image Generation with Diffusion Controller: A Simplified Unified Approach
    5 Min Read
    Effortless Long-Form Video Creation: Automating Coherent Content Generation
    Effortless Long-Form Video Creation: Automating Coherent Content Generation
    5 Min Read
  • Guides
    GuidesShow More
    Your Comprehensive Guide to Practical Constraint Decoding: Basics and Applications
    Your Comprehensive Guide to Practical Constraint Decoding: Basics and Applications
    6 Min Read
    KDnuggets Weekly Data Science News Roundup: Highlights from July 20, 2026
    KDnuggets Weekly Data Science News Roundup: Highlights from July 20, 2026
    4 Min Read
    Unlock Your AI Potential with Kaggle and Google’s Free 5-Day Agentic AI Course
    Unlock Your AI Potential with Kaggle and Google’s Free 5-Day Agentic AI Course
    6 Min Read
    Top 5 High-Performance MCP Servers for Optimal Agentic Development
    Top 5 High-Performance MCP Servers for Optimal Agentic Development
    6 Min Read
    Top 5 Free Resources for Understanding Agentic AI: Unlock Your Knowledge
    Top 5 Free Resources for Understanding Agentic AI: Unlock Your Knowledge
    6 Min Read
  • Tools
    ToolsShow More
    Create Local AI Applications Using C++ and NVIDIA TensorRT RTX Samples
    Create Local AI Applications Using C++ and NVIDIA TensorRT RTX Samples
    5 Min Read
    Unlock Near-Astra Intelligence in Your Daily Work with GPT-6.1 Sol on Amazon Bedrock
    Unlock Near-Astra Intelligence in Your Daily Work with GPT-6.1 Sol on Amazon Bedrock
    6 Min Read
    Reproducible Benchmark Results: How UK AISI and EvalEval Are Leading the Way
    Reproducible Benchmark Results: How UK AISI and EvalEval Are Leading the Way
    6 Min Read
    Hugging Face Welcomes Jun Kim, oMLX Creator and Maintainer, to Boost the MLX Community
    Hugging Face Welcomes Jun Kim, oMLX Creator and Maintainer, to Boost the MLX Community
    4 Min Read
    AWS Crowned Leader in The Forrester Wave: AI Infrastructure Solutions, Q4 2025 Report
    AWS Crowned Leader in The Forrester Wave: AI Infrastructure Solutions, Q4 2025 Report
    5 Min Read
  • Events
    EventsShow More
    Boosting Everyday Courage in Educational Leaders: A Guide to Choosing Confidence
    Boosting Everyday Courage in Educational Leaders: A Guide to Choosing Confidence
    5 Min Read
    Boosting OpenAI’s GPT-6 Astra Performance: The Role of NVIDIA GPUs in Accelerating AI Technology
    Boosting OpenAI’s GPT-6 Astra Performance: The Role of NVIDIA GPUs in Accelerating AI Technology
    4 Min Read
    Jensen Huang at Dreamforce: ‘Now We Can Know Everything and Achieve Anything’
    Jensen Huang at Dreamforce: ‘Now We Can Know Everything and Achieve Anything’
    5 Min Read
    Essential Strategies for Preparing Students for a Career in Quantum Computing
    Essential Strategies for Preparing Students for a Career in Quantum Computing
    5 Min Read
    Skild AI Leverages NVIDIA’s Physical AI to Enable Robots to Learn New Tasks from Just One Video
    Skild AI Leverages NVIDIA’s Physical AI to Enable Robots to Learn New Tasks from Just One Video
    6 Min Read
  • Ethics
    EthicsShow More
    Understanding Withholding Delay: A Welfare Model for Open-Weight AI Releases in Asymmetric Proliferation
    Understanding Withholding Delay: A Welfare Model for Open-Weight AI Releases in Asymmetric Proliferation
    6 Min Read
    Exploring Elon Musk’s Massive Midterm Election Spending Surge
    Exploring Elon Musk’s Massive Midterm Election Spending Surge
    5 Min Read
    OpenAI’s Mathematical Findings Raise Concerns Among Experts: What You Need to Know
    OpenAI’s Mathematical Findings Raise Concerns Among Experts: What You Need to Know
    4 Min Read
    Australia’s Proposed Laws: Strengthening Privacy Regulations for Chatbots – Key Details Needed for Success
    Australia’s Proposed Laws: Strengthening Privacy Regulations for Chatbots – Key Details Needed for Success
    6 Min Read
    Boost Your Work Efficiency with AI: Embrace Constructive Disagreement
    Boost Your Work Efficiency with AI: Embrace Constructive Disagreement
    6 Min Read
  • Comparisons
    ComparisonsShow More
    InternBootcamp: Enhancing LLM Reasoning Through Verifiable Task Scaling Techniques
    InternBootcamp: Enhancing LLM Reasoning Through Verifiable Task Scaling Techniques
    4 Min Read
    Enhancing Anomaly Detection in Collider Experiments through Contrastive Learning for Better Interpretability
    Enhancing Anomaly Detection in Collider Experiments through Contrastive Learning for Better Interpretability
    6 Min Read
    Exploring the Impact of Quantization on Self-Explanations in Large Language Models: Can LLMs Explain Themselves?
    Exploring the Impact of Quantization on Self-Explanations in Large Language Models: Can LLMs Explain Themselves?
    5 Min Read
    CytoNet: A Foundation Model for Understanding the Human Cerebral Cortex at Cellular Resolution
    CytoNet: A Foundation Model for Understanding the Human Cerebral Cortex at Cellular Resolution
    5 Min Read
    Optimizing Nonconvex-Nonconcave Min-Max Problems with a Limited Maximization Domain: Insights from [2110.03950]
    Optimizing Nonconvex-Nonconcave Min-Max Problems with a Limited Maximization Domain: Insights from [2110.03950]
    5 Min Read
Search
  • Privacy Policy
  • Terms of Service
  • Contact Us
  • FAQ / Help Center
  • Advertise With Us
  • Latest News
  • Model Comparisons
  • Tutorials & Guides
  • Open-Source Tools
  • Community Events
© 2025 AI Model Kit. All Rights Reserved.
Reading: Recognizing Toxicity: Understanding Span and Target in Chemical Safety
Share
Notification Show More
Font ResizerAa
AIModelKitAIModelKit
Font ResizerAa
  • 🏠
  • 🚀
  • 📰
  • 💡
  • 📚
  • ⭐
Search
  • Home
  • News
  • Models
  • Guides
  • Tools
  • Ethics
  • Events
  • Comparisons
Follow US
  • Latest News
  • Model Comparisons
  • Tutorials & Guides
  • Open-Source Tools
  • Community Events
© 2025 AI Model Kit. All Rights Reserved.
AIModelKit > Comparisons > Recognizing Toxicity: Understanding Span and Target in Chemical Safety
Comparisons

Recognizing Toxicity: Understanding Span and Target in Chemical Safety

aimodelkit
Last updated: January 7, 2026 8:15 pm
aimodelkit
Share
Recognizing Toxicity: Understanding Span and Target in Chemical Safety
SHARE
[Submitted on 2 Jun 2025 (v1), last revised 5 Jan 2026 (this version, v2)]

View a PDF of the paper titled Something Just Like TRuST: Toxicity Recognition of Span and Target, by Berk Atil and two other authors.

View PDF | HTML (experimental)

Abstract: Toxic language includes content that is offensive, abusive, or that promotes harm. Progress in preventing toxic output from large language models (LLMs) is hampered by inconsistent definitions of toxicity. We introduce TRuST, a large-scale dataset that unifies and expands prior resources through a carefully synthesized definition of toxicity and corresponding annotation scheme. It consists of ~300k annotations, with high-quality human annotation on ~11k. To ensure high quality, we designed a rigorous, multi-stage human annotation process and evaluated the diversity of the annotators. Then we benchmarked state-of-the-art LLMs and pre-trained models on three tasks: toxicity detection, identification of the target group, and of toxic words. Our results indicate that fine-tuned PLMs outperform LLMs on the three tasks, and that current reasoning models do not reliably improve performance. TRuST constitutes one of the most comprehensive resources for evaluating and mitigating LLM toxicity and other research in socially-aware and safer language technologies.

Submission History

From: Berk Atil [view email]

[v1] Mon, 2 Jun 2025 23:48:16 UTC (1,094 KB)
[v2] Mon, 5 Jan 2026 21:38:57 UTC (1,098 KB)

—

### Understanding Toxicity in Language Models

Toxic language can take many forms, often presenting as offensive, abusive, or harmful content. In today’s digital landscape, large language models (LLMs) are becoming increasingly integrated into everyday applications, from customer service to content creation. However, the challenge of ensuring these models do not propagate toxic language remains a pressing issue. This calls for a comprehensive understanding of toxicity and how to effectively manage it within LLMs.

### The TRuST Dataset: A New Benchmark

The introduction of the TRuST dataset marks a significant advancement in the field of toxicity recognition. This large-scale dataset consolidates and expands upon previous resources, creating a comprehensive framework for defining toxicity. With approximately 300,000 annotations, of which around 11,000 have undergone high-quality human annotation, TRuST aims to provide a robust foundation for future research. The dataset is instrumental for developers and researchers seeking to refine the capabilities of LLMs by providing a clear understanding of what constitutes toxic language.

More Read

Evaluating Robustness, Privacy, and Fairness in Federated Learning Combined with Foundation Models
Evaluating Robustness, Privacy, and Fairness in Federated Learning Combined with Foundation Models
Comprehensive Behavioral Testing of Large Language Models in Healthcare
Optimal Timing for Reviewing: Utilizing Spaced Repetition in Continuous Pre-Training of Language Models
Discover the 2025 QCon AI New York Schedule: Key Highlights on Practical Enterprise AI
Adaptive Video Streaming: Leveraging Semantic-Aware Latent Diffusion Models for Enhanced Wireless Network Performance

### Multi-Stage Annotation Process

The quality of any dataset is paramount, and TRuST has been developed through a meticulous multi-stage human annotation process. This was designed to not only ensure the accuracy of the toxicity classifications but also to evaluate the diversity of the annotators. Diversity among annotators helps to reduce bias in the dataset, ensuring it reflects a variety of perspectives and cultural contexts. This thoughtful approach to annotation signifies a leap forward in addressing the complexities surrounding toxic language.

### Benchmarking LLMs and Pre-trained Models

The efficacy of any dataset can be evaluated through benchmarking against existing models. In the case of TRuST, state-of-the-art LLMs and pre-trained models were assessed on three primary tasks: toxicity detection, identification of the target group, and pinpointing toxic words. The findings revealed that fine-tuned pre-trained language models (PLMs) significantly outperform LLMs in these tasks. This insight is crucial for developers aiming to build safer language technologies, as it informs the selection of models based on specific functionalities.

### Addressing Current Limitations

One of the key discoveries in the TRuST study is that the current reasoning models do not consistently enhance performance in toxicity detection. This indicates a critical area for future research, highlighting the need for ongoing development to improve the capabilities of these models. It brings to light the importance of understanding the limitations of current technologies and the necessity for continual innovation in the field of natural language processing.

### Implications for Safer Language Technologies

TRuST stands out as one of the most comprehensive resources available for evaluating and mitigating toxicity in large language models. It paves the way for further research into socially-aware language technologies, enabling developers and researchers to create applications that are not only effective but also responsible. By utilizing the TRuST dataset, stakeholders can contribute to building a digital landscape that prioritizes safety, inclusivity, and respect, making strides towards a more harmonious online environment.

### Future Directions in Toxicity Research

As the discourse surrounding toxic language continues to evolve, the need for innovative solutions remains significant. Future research should focus on enhancing the methodologies used to identify and mitigate toxicity, as well as expanding datasets like TRuST to encapsulate a broader spectrum of language nuances. By fostering collaboration within the research community, we can collectively work towards developing more sophisticated tools that effectively curb the spread of toxic language in digital communication.

—

This article serves as a detailed exploration of the TRuST dataset and its implications in the fight against toxic language in AI. By focusing on critical aspects such as annotation quality, benchmarking, and future directions for research, it provides a well-rounded understanding of the ongoing challenges and advancements in this essential area of study.

Inspired by: Source

Enhancing Multimodal Fact-Checking with an Agent-Based Approach: Insights from Study [2512.22933]
Exploring the Resilience of Knowledge Tracing Models Against Student Concept Drift: Insights from Research [2511.00704]
Cloudflare Enhances D1 Database with Global Read Replication Features
Enhanced NovaSAR Dataset for Automated Ship Target Recognition
Exploring Sentence Transformers on the Hugging Face Hub: A Comprehensive Guide

Sign Up For Daily Newsletter

Get AI news first! Join our newsletter for fresh updates on open-source models.

By signing up, you agree to our Terms of Use and acknowledge the data practices in our Privacy Policy. You may unsubscribe at any time.
Share This Article
Facebook Copy Link Print
Previous Article Ultimate Guide to Converting Bytes to Strings in Python: Take the Quiz – Real Python Ultimate Guide to Converting Bytes to Strings in Python: Take the Quiz – Real Python
Next Article Dell Acknowledges Consumer Disinterest in AI-Driven PCs Dell Acknowledges Consumer Disinterest in AI-Driven PCs

Stay Connected

XFollow
PinterestPin
TelegramFollow
LinkedInFollow

							banner							
							banner
Explore Top AI Tools Instantly
Discover, compare, and choose the best AI tools in one place. Easy search, real-time updates, and expert-picked solutions.
Browse AI Tools

Latest News

Understanding Withholding Delay: A Welfare Model for Open-Weight AI Releases in Asymmetric Proliferation
Understanding Withholding Delay: A Welfare Model for Open-Weight AI Releases in Asymmetric Proliferation
Ethics
Exploring Elon Musk’s Massive Midterm Election Spending Surge
Exploring Elon Musk’s Massive Midterm Election Spending Surge
Ethics
Boosting Everyday Courage in Educational Leaders: A Guide to Choosing Confidence
Boosting Everyday Courage in Educational Leaders: A Guide to Choosing Confidence
Events
Unlocking Efficient Autoregressive Video Generation with SemanTok: Predictable Semantic Tokens by Stability AI
Unlocking Efficient Autoregressive Video Generation with SemanTok: Predictable Semantic Tokens by Stability AI
Open-Source Models
//

Leading global tech insights for 20M+ innovators

Quick Link

  • Latest News
  • Model Comparisons
  • Tutorials & Guides
  • Open-Source Tools
  • Community Events

Support

  • Privacy Policy
  • Terms of Service
  • Contact Us
  • FAQ / Help Center
  • Advertise With Us

Sign Up for Our Newsletter

Get AI news first! Join our newsletter for fresh updates on open-source models.

AIModelKitAIModelKit
Follow US
© 2025 AI Model Kit. All Rights Reserved.
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?