asiai bronze
asiai.dev
“Apple Silicon LLM inference benchmark and monitoring agent. Exposes 11 read-only tools and 3 resources over the Model Context Protocol (MCP) to detect installed inference engines, benchmark local models, and recommend configurations by hardware. Runs locally (stdio) or over SSE/streamable-HTTP.
a2a https://asiai.dev talk to it https://asiai.dev/.well-known/agent-card.json its cardwe checked this the operator says this
Verified by agenttru.st
Everything here is a check agenttru.st performed itself. Assurance, protocol, hosting and freshness are in the card above and are not repeated.
- Certificate
-
Issued by Let's Encrypt
domain-validated
Valid until 1 Dec 2026.Control of the hostname was checked; nothing about who operates it.
- DANE / TLSA
- Not verified (TLSA query returned RCodeNameError)
- Discovery
- Well-known document
- AI use policy
-
What this site's
robots.txtsays about how AI may use its content. Recorded as the operator wrote it, not enforced β these are preferences about use, not access, and agenttru.st only reads the agent's own discovery documents. - First seen
- 2 Sep 2026
View verification details
- Assurance
- bronze Bronze β agent card fetched over HTTPS with a valid certificate
- Protocols
- A2A verified by handshake or card fetch, not merely advertised
- Hosted in
- πΊπΈ US Β· Fastly, Inc. (AS54113)
- Last checked
- 4d ago
What this agent says it can do
Declared in the agent's own card. agenttru.st has not tested whether it completes any of these tasks β the operator of asiai.dev controls every word below.
Check Inference Health
Quick health check of all local LLM inference engines. Returns ok/degraded/error, memory pressure, thermal state, GPU. Responds in <500ms.
- Is local LLM inference available right now?
List Loaded Models
List all models currently loaded across inference engines (VRAM, quantization, context length).
- What models are loaded right now?
Detect Inference Engines
Auto-detect running LLM inference engines (Ollama, LM Studio, mlx-lm, llama.cpp, vLLM-MLX, Exo, TurboQuant).
- Which inference engines are installed on this Mac?
Run Inference Benchmark
Benchmark a local model's performance (tok/s, TTFT, VRAM, power) with statistical rigour (CI 95%, P50/P90/P99). Supports multi-engine and cross-model comparison.
- Benchmark Qwen 3.6 on Ollama NVFP4
- Compare Qwen 3.5 vs 3.6 on this Mac
Recommend Engine and Model
Hardware-aware engine+model recommendations optimized for throughput, latency, or power efficiency.
- What's the fastest engine for my Mac?
- Which model fits my RAM?
Compare Engines
Side-by-side comparison of inference engines or models from benchmark history.
- Compare Ollama MLX vs LM Studio for Qwen 3.6
Full Inference Snapshot
Complete system + inference state: CPU load, memory, thermal, GPU, engines status, loaded models, recent activity.
- Give me a full status report of local inference
Run Diagnostics
Comprehensive diagnostic checks: Apple Silicon compat, engines health, DB integrity, daemon status, alerting config.
- Diagnose why inference is failing
Technical agent card
Copied from the agent's card. The operator controls these values; agenttru.st has not verified them.
- Provider
- asiai (Jean-Marc Nahlovsky / druide67) β what this agent says about itself; other agents claiming the same provider are not thereby related
- Protocol
- a2a
- Version
- 1.6.0
- Card completeness
- complete all eight fields required by a2a.proto v1.0
View all card details
- Capabilities
- pushNotifications stateTransitionHistory streaming
- Agent card
- https://asiai.dev/.well-known/agent-card.json
Operate this agent and would rather not be listed? Request removal.