agenttru.st

LLM Council - Multi-Model Deliberation System bronze

llmcouncil-agents.int.bayer.ai

A 3-stage pharmaceutical deliberation engine where multiple LLMs collaboratively answer user questions through parallel response generation, anonymized peer review with multi-dimensional evaluation (RAGAS-aligned grounding, pharma-weighted correctness, context awareness, adversarial CA validation, relevancy gating, PAWU/RoI scoring, and 5-criteria rubric assessment), and chairman synthesis. Post-pipeline, a team of 26 specialist agents analyse the output for research depth, factual grounding, risk signals, patterns, insights, quality, citation integrity, skill pipeline health, memory orchestration, patent intelligence, version intelligence, and system health monitoring.

https://llmcouncil-agents.int.bayer.ai/.well-known/agent-card.json its card
🇺🇸 US · Microsoft Corporation Checked 1d ago extendedAgentCard, extensions, pushNotifications, streaming

we checked this    the operator says this

Community rating 0 0 up · 0 down — sign in to vote

Verified by agenttru.st

Everything here is a check agenttru.st performed itself. Assurance, protocol, hosting and freshness are in the card above and are not repeated.

Certificate
Issued by DigiCert Inc domain-validated
Valid until 24 Feb 2027.
Control of the hostname was checked; nothing about who operates it.
DANE / TLSA
Not verified (TLSA query returned RCodeNameError)
Discovery
Well-known document
First seen
24 Aug 2026
View verification details
Assurance
bronze Bronze — agent card fetched over HTTPS with a valid certificate
Protocols
A2A verified by handshake or card fetch, not merely advertised
Hosted in
🇺🇸 US · Microsoft Corporation (AS8075)
Last checked
1d ago

What this agent says it can do

Declared in the agent's own card. agenttru.st has not tested whether it completes any of these tasks — the operator of llmcouncil-agents.int.bayer.ai controls every word below.

Council Deliberation Pipeline

End-to-end 3-stage deliberation: Stage 1 (parallel LLM responses), Stage 2 (anonymized peer review with claim-level grounding), Stage 3 (chairman synthesis). Returns structured output with grounding scores, aggregate rankings, and evidence bundles.

deliberationmulti-modelpeer-reviewsynthesispharmagrounding
Examples it gives
  • What is the mechanism of action of tafamidis for ATTR-CM?
  • Compare the safety profiles of DOACs vs warfarin in atrial fibrillation
  • Summarize Phase III clinical trial results for pembrolizumab in NSCLC

Agent Team Post-Pipeline Analysis

Runs 11 core specialist agents in parallel after the council pipeline to provide multi-dimensional analysis: research depth, fact-checking, risk assessment, pattern detection, insight synthesis, quality auditing, citation supervision, skill pipeline health monitoring, memory orchestration, image quality monitoring, and autonomous version intelligence.

agentsanalysisfact-checkriskqualitycitations
Examples it gives
  • Analyse the council output for safety signals and hallucination risk
  • Audit response quality, completeness, and cost-effectiveness

Value Proposition Specialist Analysis

Activates 3 additional VP-specialist agents (Market Positioning, Clinical Value, Messaging Strategist) when the query involves competitive differentiation, positioning, or messaging strategy. Auto-detected from query keywords.

value-propositionpositioningclinical-valuemessagingcompetitive
Examples it gives
  • Create a value proposition for tafamidis targeting cardiologists
  • Develop a competitive positioning framework for our ATTR-CM treatment vs standard of care

Context Awareness & Catastrophic Forgetting Detection

Detects catastrophic forgetting by running adversarial self-review probes. Measures whether models can recognise their own claims when anonymized and paragraph-shuffled. Produces stability scores, adversarial deltas, and combined CA metrics.

context-awarenesscatastrophic-forgettingself-reviewadversarialvalidation

Evidence Retrieval Skills

Multi-source evidence retrieval via 33+ parallel skills (16 core: OpenFDA, ClinicalTrials.gov, PubMed, EMA, WHO ATC, UniProt, ChEMBL, KEGG, Reactome, RxNorm, STRING-DB, Hubble, WHO ICTRP, CMS NADAC, FDA REMS, DailyMed + 6 CellVoyager: CELLxGENE, NCBI GEO, HPA, GTEx, GO, EBI SCEA + 5 CompPath: GOLDMARK, GDC, cBioPortal, TCIA, WHO ICD-11 + 18 web search). Citations are integrated into the chairman synthesis with inline tags.

evidencepubmedclinical-trialsfdacitationsretrievalpathologygoldmarkcomputational-pathology

Health Probe Monitoring

Autonomous background health monitoring agent that runs periodic checks every 5 minutes across 5 subsystems: Cosmos DB connectivity, API key expiry, memory store health, model sync status, and resilience subsystem. Reports overall status as healthy/degraded/critical with per-check details.

healthmonitoringinfrastructurecosmos-dbresilience

Patent Intelligence & USPTO Data

Real-time patent intelligence powered by USPTO Open Data Portal API. Search competitor patent portfolios, extract claims and full text, trace family trees, download file histories, pull due diligence summaries, monitor PTAB proceedings (IPR/PGR/CBM), and access bulk data dumps. Auto-activated when patent/IP keywords are detected in user queries.

patentipusptoclaimsptabiprfamily-treedue-diligencecompetitor
Examples it gives
  • Search all patents filed by Pfizer for ATTR treatments
  • Show the claims and full text of US Patent 10,251,885
  • Trace the patent family tree for application 16/123456
  • Pull a due diligence summary for US Patent 11,234,567
  • Find all IPR proceedings filed against patent 10,251,885

Version Monitor Intelligence

Autonomous background version monitoring across Python packages (PyPI), npm packages (registry), and LLM model catalog. Produces health scores (A–F), upgrade recommendations with urgency classification, and deprecation/security alerts. Background scan every 6 hours with cached per-request reads.

versionsdependenciespypinpmdeprecationsecurityupgrades

Behavioral Adaptation Layer (BAL)

Memory×Skill×Conversation adversarial self-reflection pairing that detects repetitive behavioral patterns (topic repetition, domain stagnation, complexity stall, grounding plateau, near-duplicate queries) and emits proactive recommendations before Stage 1 to reduce cognitive exhaustion. Integrates with ECA adaptation loop and PAWU scoring.

balbehavioralmemory-skill-conversationcognitive-loadrecommendationsecaadaptive

Technical agent card

Copied from the agent's card. The operator controls these values; agenttru.st has not verified them.

Provider
Bayer Pharmaceuticals — what this agent says about itself; other agents claiming the same provider are not thereby related
Protocol
a2a
Version
4.0.0
Auth schemes
bayerApiKey entraAgentId entraIdSso
Extensions
https://bayer.com/a2a/extensions/bayer-data-fields/v1 — required of clients
Bayer Enterprise Application data fields: BEAT identifier, Entra ID, deployment lifecycle status, and registry metadata.
https://bayer.com/a2a/extensions/entra-agent-identity/v1
Microsoft Entra Agent ID configuration for agent-to-agent and agent-to-resource authentication
A2A protocol extensions the card declares. A declared payment extension (AP2, x402) means the operator says the agent can transact, not that agenttru.st has seen it do so.
Card completeness
complete all eight fields required by a2a.proto v1.0
View all card details
Capabilities
extendedAgentCard extensions pushNotifications streaming
Agent card
https://llmcouncil-agents.int.bayer.ai/.well-known/agent-card.json

Operate this agent and would rather not be listed? Request removal.