Anzeige
Mehr »
Mittwoch, 23.09.2026 - Börsentäglich über 12.000 News
Brandneue News liefert jetzt Teil 2 dieser Kupfer-Story
Anzeige

Indizes

Kurs

%
News
24 h / 7 T
Aufrufe
7 Tage

Aktien

Kurs

%
News
24 h / 7 T
Aufrufe
7 Tage

Xetra-Orderbuch

Fonds

Kurs

%

Devisen

Kurs

%

Rohstoffe

Kurs

%

Themen

Kurs

%

Erweiterte Suche
ACCESS Newswire
153 Leser
Artikel bewerten:
(0)

TrustScale Launches ArgusRL as Automated AI Evaluation Surpasses Human Performance

Customer calls it a "singularity moment" for reinforcement learning with human feedback, as evidence-grounded automation crosses the human-quality threshold at scale

LOS ALTOS, CA / ACCESS Newswire / September 23, 2026 / TrustScale, an AI training, evaluation and assurance company, today announced the launch of ArgusRL, an automated evaluation and reinforcement feedback platform that outperformed a customer's highest qualified human evaluators in a production deployment, including identifying errors human reviewers missed. In testing, more than 95% of ArgusRL's automated evaluations were accepted by the customer without correction.

"One of our early customers, a leading global technology company, told us we've reached a singularity moment for reinforcement learning with human feedback, where evidence-grounded automation crosses the human-quality threshold at scale," said Lawrence Snapp, CEO of TrustScale. "This fundamentally changes the economics of AI training and reinforcement learning. AI makers and deployers no longer have to choose between the scale of automation and the quality of human evaluation. They can have both, grounded in deterministic evidence rather than another probabilistic AI opinion."

Human feedback has long been the gold standard for evaluating and improving AI models through reinforcement learning. As AI development accelerates, model makers are increasingly automating that process with AI judges and other model-based evaluation systems. ArgusRL takes a different approach, using empirical evidence and deterministic verification to evaluate AI-generated responses and generate structured feedback that can be used to continuously improve model performance.

In a production deployment with a leading global technology company, ArgusRL's automated evaluation delivered better results than the customer's human annotators and identified mistakes the human reviewers had missed.

"We've worked with a range of partners and approaches to improve the quality of reinforcement learning and model evaluation, and ArgusRL has consistently stood out for the quality and accuracy of its prompt and response review," Former Apple and Amazon AGI Leader. "Its ability to identify errors missed during human review is particularly compelling, demonstrating the potential for deterministic automation to improve both the quality and scale of AI evaluation."

Unlike AI-Judge approaches that rely solely on probabilistic AI to evaluate another probabilistic system, TrustScale's patent-pending ArgusRL technology grounds its evaluations in retrieved external evidence. The platform analyzes each prompt and response, breaks responses into individual claims and searches multiple data sources for supporting or contradictory evidence. It then returns structured deterministic verdicts with citations and confidence scores.

ArgusRL also evaluates the quality of the original query and overall response and identifies cases that warrant human review. With ArgusRL, contradicted claims, claims without sufficient evidence and other flagged responses can be routed to human annotators, allowing people to focus on the cases where human judgment adds the greatest value rather than manually evaluating every response.

Because ArgusRL continuously evaluates outputs after deployment, its reinforcement feedback can incorporate current evidence and information that may not have been available during a model's initial training.

"The implications go well beyond accuracy," said Snapp. "AI companies spend billions of dollars each year on the data, human evaluation and infrastructure required to train and improve models. If AI makers can automate more of the reinforcement feedback process without sacrificing quality, they can improve models faster and at lower cost while reserving human expertise for the cases that actually require it."

ArgusRL is built on the same evidence-based TrustScale Engine that powers Argus, the company's AI assurance platform for detecting and correcting hallucinations at the point of use. ArgusRL takes that evidence-based approach upstream, giving AI makers and developers structured feedback they can incorporate into model training, fine-tuning pipeline, and continuous improvement.

ArgusRL operates as an API-backed evaluation service and supports multiple languages, locales and input formats. It can be integrated with existing model development, evaluation, and annotation workflows and returns claim-level verdicts, supporting evidence, citations, and structured results for downstream use.

ArgusRL is available today through the AWS Marketplace and directly through TrustScale. To learn more or request a demonstration, visit https://trustscale.ai/en/argusrl.

About TrustScale

TrustScale is an AI training, evaluation and assurance company helping organizations create, shape and use artificial intelligence with greater confidence and control. Built on more than 20 years of experience in AI data across 200-plus languages, TrustScale develops independent technologies that detect AI mistakes, evaluate claims against deterministic empirical evidence and keep people at the center of consequential decisions. Its Argus suite spans the AI lifecycle, from real-time hallucination detection and evidence-based correction at the point of use, to automated evaluation and reinforced feedback for model training and continuous improvement. Learn more at TrustScale.ai.

Media contact:

Songue PR for TrustScale
trustscale@songuepr.com

SOURCE: TrustScale



View the original press release on ACCESS Newswire:
https://www.accessnewswire.com/newsroom/en/computers-technology-and-internet/trustscale-launches-argusrl-as-automated-ai-evaluation-surpasses-1225554

© 2026 ACCESS Newswire
KI-Euphorie kippt - Bei diesen 5 Aktien droht der Crash!
Drei Jahre lang kannten KI-Aktien fast nur eine Richtung: nach oben. Billionenschwere Investitionspläne von Alphabet, Amazon, Meta und Microsoft haben Halbleiter- und Infrastrukturwerte auf immer neue Höhen getrieben. Doch jetzt bekommt die Erfolgsstory gefährliche Risse.

Steigende Anleiherenditen verteuern die Finanzierung, während die gewaltigen KI-Ausgaben zunehmend nicht mehr aus den laufenden Cashflows bezahlt werden können. Gleichzeitig zeigen günstigere chinesische Modelle, dass leistungsfähige KI womöglich mit deutlich weniger Rechenleistung auskommt. Damit wächst die Gefahr, dass heute für Milliarden errichtete Kapazitäten morgen nicht die erhofften Renditen liefern.

Für Anleger könnte das zum Problem werden. Denn treffen steigende Finanzierungskosten auf Überkapazitäten und enttäuschende Cashflows, geraten gerade hoch bewertete KI-Profiteure schnell unter Druck. Aus den größten Gewinnern der vergangenen Jahre könnten so die größten Verlierer der nächsten Korrektur werden.

In unserem aktuellen Spezialreport zeigen wir 5 Aktien, bei denen das Chance-Risiko-Verhältnis jetzt besonders gefährlich erscheint – und bei denen Anleger genauer hinschauen sollten.

Jetzt den kostenlosen Report sichern – bevor die KI-Euphorie ihren nächsten Realitätstest erlebt!
Werbehinweise: Die Billigung des Basisprospekts durch die BaFin ist nicht als ihre Befürwortung der angebotenen Wertpapiere zu verstehen. Wir empfehlen Interessenten und potenziellen Anlegern den Basisprospekt und die Endgültigen Bedingungen zu lesen, bevor sie eine Anlageentscheidung treffen, um sich möglichst umfassend zu informieren, insbesondere über die potenziellen Risiken und Chancen des Wertpapiers. Sie sind im Begriff, ein Produkt zu erwerben, das nicht einfach ist und schwer zu verstehen sein kann.