# Trustworthy AI

An artificial intelligence discipline focused on incorporating trustworthiness considerations throughout AI system design, development, deployment, use, and evaluation.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## Using deterministic systems and provenance to make AI-generated financial analysis verifiable

DevFeed: [Using deterministic systems and provenance to make AI-generated financial analysis verifiable](<https://devfeed.tech/articles/this-shit-is-hard-getting-ai-to-prove-where-a-number-came-from-13279.md>)

Original publisher: [Read original article](<https://www.chainguard.dev/unchained/this-shit-is-hard-getting-ai-to-prove-where-a-number-came-from>)

Published: 2026-08-25T00:00:00Z

Content type: article

Language: en

Sources: [Chainguard: Unchained](<https://devfeed.tech/sources/chainguard-unchained.md>)

Topics: [Trustworthy AI](<https://devfeed.tech/topics/trustworthy-ai.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [AI Platform](<https://devfeed.tech/topics/ai-platform.md>), [Large Language Model](<https://devfeed.tech/topics/llm.md>), [Code](<https://devfeed.tech/topics/code.md>), [Finance](<https://devfeed.tech/topics/finance.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-platform](<https://devfeed.tech/tags/ai-platform.md>), [financial](<https://devfeed.tech/tags/financial.md>), [models](<https://devfeed.tech/tags/models.md>), [provenance](<https://devfeed.tech/tags/provenance.md>), [trustworthy-ai](<https://devfeed.tech/tags/trustworthy-ai.md>), [verification](<https://devfeed.tech/tags/verification.md>)

### AI overview

Kepler describes a model-agnostic approach to trustworthy AI that combines language models with deterministic tools for retrieval, computation, provenance, and traceability. The article focuses on making financial analysis outputs verifiable by linking numbers to their sources, formulas, or computations.

### Source excerpt

Trustworthy AI takes more than a powerful model. See how Kepler uses deterministic systems and provenance to make financial analysis verifiable.

## Show, Don't Tell: What Evo Continuous Offensive Security Found in a Real Enterprise SaaS

DevFeed: [Show, Don't Tell: What Evo Continuous Offensive Security Found in a Real Enterprise SaaS](<https://devfeed.tech/articles/show-don-t-tell-what-evo-continuous-offensive-security-found-in-a-real-enterprise-saas-8243.md>)

Original publisher: [Read original article](<https://snyk.io/blog/what-evo-cos-found-real-enterprise-saas/>)

Author: Nuno Loureiro; Luis Grangeia

Published: 2026-08-10T00:00:00Z

Content type: article

Language: en

Sources: [Blog RSS Feed | Snyk](<https://devfeed.tech/sources/blog-rss-feed-snyk.md>)

Topics: [Security](<https://devfeed.tech/topics/security.md>), [Vulnerabilities](<https://devfeed.tech/topics/vulnerabilities.md>), [Authorization](<https://devfeed.tech/topics/authorization.md>), [Software as a service](<https://devfeed.tech/topics/saas.md>), [Microservice](<https://devfeed.tech/topics/microservice.md>), [Microservices](<https://devfeed.tech/topics/microservices.md>), [Front end](<https://devfeed.tech/topics/frontend.md>), [API](<https://devfeed.tech/topics/api.md>), [App](<https://devfeed.tech/topics/app.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Trustworthy AI](<https://devfeed.tech/topics/trustworthy-ai.md>)

Tags: [agent](<https://devfeed.tech/tags/agent.md>), [ai](<https://devfeed.tech/tags/ai.md>), [api](<https://devfeed.tech/tags/api.md>), [application-security](<https://devfeed.tech/tags/application-security.md>), [applications](<https://devfeed.tech/tags/applications.md>), [attacks](<https://devfeed.tech/tags/attacks.md>), [autonomous](<https://devfeed.tech/tags/autonomous.md>), [awareness](<https://devfeed.tech/tags/awareness.md>), [blog](<https://devfeed.tech/tags/blog.md>), [customer](<https://devfeed.tech/tags/customer.md>), [customer-featured](<https://devfeed.tech/tags/customer-featured.md>), [cybersecurity](<https://devfeed.tech/tags/cybersecurity.md>), [devops](<https://devfeed.tech/tags/devops.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [frontend](<https://devfeed.tech/tags/frontend.md>), [interest](<https://devfeed.tech/tags/interest.md>), [microservices](<https://devfeed.tech/tags/microservices.md>), [saas](<https://devfeed.tech/tags/saas.md>), [security](<https://devfeed.tech/tags/security.md>), [vulnerabilities](<https://devfeed.tech/tags/vulnerabilities.md>), [vulnerability-insights](<https://devfeed.tech/tags/vulnerability-insights.md>), [web](<https://devfeed.tech/tags/web.md>)

### AI overview

Evo Continuous Offensive Security (COS) is presented as an autonomous offensive-security system combining AI pentesting, agent red teaming, and dynamic testing. The article reports a real assessment of a multi-tenant enterprise SaaS application that uncovered 33 confirmed vulnerabilities, including tenant-wide compromise and critical authorization flaws.

### Source excerpt

A real Evo Continuous Offensive Security assessment uncovered 33 confirmed vulnerabilities in a multi-tenant enterprise SaaS, including tenant-wide compromise and critical authorization flaws.

## LLM reasoning and agentic safety at ICML 2026

DevFeed: [LLM reasoning and agentic safety at ICML 2026](<https://devfeed.tech/articles/llm-reasoning-and-agentic-safety-at-icml-2026-22575.md>)

Original publisher: [Read original article](<https://medium.com/capital-one-tech/llm-reasoning-and-agentic-safety-at-icml-2026-55f341e21caa?source=rss----3db3a67cb648---4>)

Author: Capital One Tech

Published: 2026-07-02T14:14:40Z

Content type: article

Language: en

Sources: [Capital One Tech](<https://devfeed.tech/sources/capital-one-tech.md>)

Topics: [Large Language Model](<https://devfeed.tech/topics/llm.md>), [Fine-tuning](<https://devfeed.tech/topics/fine-tuning.md>), [Trustworthy AI](<https://devfeed.tech/topics/trustworthy-ai.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>)

Tags: [2026](<https://devfeed.tech/tags/2026.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-research](<https://devfeed.tech/tags/ai-research.md>), [fine-tuning](<https://devfeed.tech/tags/fine-tuning.md>), [icml](<https://devfeed.tech/tags/icml.md>), [icml-2026](<https://devfeed.tech/tags/icml-2026.md>), [llm](<https://devfeed.tech/tags/llm.md>), [llm-reasoning](<https://devfeed.tech/tags/llm-reasoning.md>), [reasoning](<https://devfeed.tech/tags/reasoning.md>), [research](<https://devfeed.tech/tags/research.md>), [safety](<https://devfeed.tech/tags/safety.md>), [science](<https://devfeed.tech/tags/science.md>), [trustworthy-ai](<https://devfeed.tech/tags/trustworthy-ai.md>)

### AI overview

Capital One presents research for ICML 2026 on critique-guided distillation for robust LLM reasoning and on safety risks in multi-turn tool-using agents. The article says its Critique-Guided Distillation framework trains models to refine flawed responses using teacher critiques, and reports higher mathematical-reasoning benchmark performance than critique fine-tuning and standard distillation, including a 7% average improvement and gains of up to 15.0% on AMC23 and 12.2% on MATH-500.

### Source excerpt

Explore our latest research in critique-guided distillation and multi-turn agent uncertainty in Seoul.Explore our latest research in critique-guided distillation and multi-turn agent uncertainty in Seoul. Capital One technologists are excited to participate in the 43rd International Conference on Machine Learning (ICML) taking place at the COEX Convention & Exhibition Center in Seoul, South Korea, July 6-11, 2026. As a premier global venue for machine learning research, ICML provides an essential forum for exploring foundational advancements, algorithmic innovations and cutting-edge deep learning systems. Capital One is excited to share advancements in large language model (LLM) scaling efficiencies, multi-turn tool-using agent safety and the development of robust, trustworthy AI frameworks. This work delivers the underlying engineering and algorithmic improvements crucial for deploying the next generation of safe financial technologies. Main conference research: Robust reasoning and agentic risk The following research, accepted to the ICML Main Conference, pushes the boundaries of how models self-correct, how trajectory-level risks can be proactively flagged, and how multi-turn agent interactions maintain reliable execution. This section features work led by Capital One researchers alongside deep collaborations with academic partners. Critique-Guided Distillation for Robust Reasoning via Refinement Capital One Authors: Berkcan Kapusuzoglu, Supriyo Chakraborty, Michael Lee, Sambit Sahu Supervised fine-tuning with expert demonstrations often produces models that imitate outputs without internalizing the reasoning processes needed for robust generalization. While critique-based approaches show promise, training models to generate critiques directly, such as Critique Fine-Tuning (CFT), can lead to output-format drift and degradation of general capabilities. We propose Critique-Guided Distillation (CGD), a training framework that decouples critique consumption from crit

## Supporting Europe's work in ensuring a trustworthy AI ecosystem

DevFeed: [Supporting Europe's work in ensuring a trustworthy AI ecosystem](<https://devfeed.tech/articles/supporting-europe-s-work-in-ensuring-a-trustworthy-ai-ecosystem-6671.md>)

Original publisher: [Read original article](<https://openai.com/index/supporting-eu-trustworthy-ai-ecosystem>)

Published: 2026-06-11T00:00:00Z

Content type: release

Language: en

Sources: [OpenAI News](<https://devfeed.tech/sources/openai-news.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [OpenAI](<https://devfeed.tech/topics/openai.md>), [Trustworthy AI](<https://devfeed.tech/topics/trustworthy-ai.md>), [ai-governance](<https://devfeed.tech/topics/ai-governance.md>), [Code](<https://devfeed.tech/topics/code.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-content](<https://devfeed.tech/tags/ai-content.md>), [ai-governance](<https://devfeed.tech/tags/ai-governance.md>), [ecosystem](<https://devfeed.tech/tags/ecosystem.md>), [eu](<https://devfeed.tech/tags/eu.md>), [global-affairs](<https://devfeed.tech/tags/global-affairs.md>), [openai](<https://devfeed.tech/tags/openai.md>), [trustworthy-ai](<https://devfeed.tech/tags/trustworthy-ai.md>)

### AI overview

OpenAI announces support for the EU Code of Practice on Transparency of AI-Generated Content, building on provenance standards, C2PA metadata, marking and detection methods, research, and a public verification tool. The article frames provenance as a way to provide context about content and strengthen a trustworthy AI ecosystem.

### Source excerpt

OpenAI supports the EU Code of Practice on AI content transparency, advancing provenance standards and tools to help people understand AI-generated content.

## Teaching AI to see the world more like we do

DevFeed: [Teaching AI to see the world more like we do](<https://devfeed.tech/articles/teaching-ai-to-see-the-world-more-like-we-do-6251.md>)

Original publisher: [Read original article](<https://deepmind.google/blog/teaching-ai-to-see-the-world-more-like-we-do/>)

Author: Andrew Lampinen; Klaus Greff

Published: 2025-11-11T11:49:13Z

Content type: article

Language: en

Sources: [Google DeepMind News](<https://devfeed.tech/sources/google-deepmind-news.md>)

Topics: [Human-AI evaluation](<https://devfeed.tech/topics/human-ai-evaluation.md>), [Trustworthy AI](<https://devfeed.tech/topics/trustworthy-ai.md>), [AI Models](<https://devfeed.tech/topics/ai-models.md>), [AI Research](<https://devfeed.tech/topics/ai-research.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-models](<https://devfeed.tech/tags/ai-models.md>), [model](<https://devfeed.tech/tags/model.md>), [research](<https://devfeed.tech/tags/research.md>), [science](<https://devfeed.tech/tags/science.md>), [trustworthy-ai](<https://devfeed.tech/tags/trustworthy-ai.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

A new Nature paper examines how AI vision models organize visual representations differently from humans. The researchers show that reorganizing these representations to better align with human knowledge can improve models' robustness, reliability, and ability to generalize.

### Source excerpt

Our new paper analyzes the important ways AI systems organize the visual world differently from humans.

## From Ideas to Impact: How the Bay Area Is Shaping the Future of Secure AI

DevFeed: [From Ideas to Impact: How the Bay Area Is Shaping the Future of Secure AI](<https://devfeed.tech/articles/from-ideas-to-impact-how-the-bay-area-is-shaping-the-future-of-secure-ai-7930.md>)

Original publisher: [Read original article](<https://snyk.io/blog/from-ideas-to-impact/>)

Author: Manoj Nair

Published: 2025-08-06T04:00:00Z

Content type: article

Language: en

Sources: [Blog RSS Feed | Snyk](<https://devfeed.tech/sources/blog-rss-feed-snyk.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Trustworthy AI](<https://devfeed.tech/topics/trustworthy-ai.md>), [ai security](<https://devfeed.tech/topics/ai-security.md>), [Securing AI](<https://devfeed.tech/topics/securing-ai.md>), [Generative AI](<https://devfeed.tech/topics/generative-ai.md>), [MCP Server](<https://devfeed.tech/topics/mcp-server.md>), [Responsibility & Safety](<https://devfeed.tech/topics/responsibility-safety.md>), [cursor](<https://devfeed.tech/topics/cursor.md>), [CI/CD](<https://devfeed.tech/topics/cicd.md>), [sdlc](<https://devfeed.tech/topics/sdlc.md>), [Development](<https://devfeed.tech/topics/development.md>), [AWS Transform](<https://devfeed.tech/topics/aws-transform.md>)

Tags: [agentic-ai](<https://devfeed.tech/tags/agentic-ai.md>), [agents](<https://devfeed.tech/tags/agents.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-security](<https://devfeed.tech/tags/ai-security.md>), [application-security](<https://devfeed.tech/tags/application-security.md>), [awareness](<https://devfeed.tech/tags/awareness.md>), [blog](<https://devfeed.tech/tags/blog.md>), [ci-cd](<https://devfeed.tech/tags/ci-cd.md>), [compliance](<https://devfeed.tech/tags/compliance.md>), [cursor](<https://devfeed.tech/tags/cursor.md>), [development](<https://devfeed.tech/tags/development.md>), [devsecops](<https://devfeed.tech/tags/devsecops.md>), [executive](<https://devfeed.tech/tags/executive.md>), [mcp-server](<https://devfeed.tech/tags/mcp-server.md>), [sdlc](<https://devfeed.tech/tags/sdlc.md>), [secure-by-design](<https://devfeed.tech/tags/secure-by-design.md>), [security](<https://devfeed.tech/tags/security.md>), [snyk-platform](<https://devfeed.tech/tags/snyk-platform.md>), [trust](<https://devfeed.tech/tags/trust.md>), [updates](<https://devfeed.tech/tags/updates.md>)

### AI overview

This article reports insights from Snyk's Silicon Valley Lighthouse event on building secure, trustworthy AI systems. It examines agentic applications, evolving development workflows, real-time security feedback, and a culture of shared responsibility across development, platform, and security teams.

### Source excerpt

Insights from Snyk's Silicon Valley Lighthouse event on building secure, trustworthy AI. Learn how to secure agentic apps, embrace a "secure by everyone" culture, and use guardrails for AI innovation.