Metal Mantra / AGENT INTELLIGENCEPublic evidence reportAUTOMATED SIGNAL. NOT CERTIFICATION.
PUBLIC REPORT / AGENT SIGNAL v0.2

smolagents.

Scanned repository says: 🤗 smolagents: a barebones library for agents. Agents write python code to call tools or orchestrate other agents.

Inspect GitHub source

29.7k starsApache-2.0pushed 3d ago

0community interest

Votes show community interest, not safety. They never change this score.

AUTOMATED SIGNAL71/100BBBNot security certification
START HERE

Critical findings need attention.

Critical finding · details withheld for 30 days

Inspect the findings
WHAT THIS DOES NOT ESTABLISH

Runtime safety, real-world agent capability, fitness for your use case or independent certification.

Scan scope48 / 48Complete eligible-file selection · not the entire repository
Recorded findings7Heuristic matches requiring human review
SnapshotOct 9, 2026UTC · engine 0.2.0
DNS at scan timeUnverifiedDomain control is not security certification
01 / KNOW THE REPOSITORY

What the source
actually tells us.

Structured, cited information from the scanned archive. Repository-authored claims are labeled and never treated as verified capabilities.

INSPECTED SOURCE48 / 48Complete eligible selection
READMEReadRepository-authored information
MANIFESTS2LICENSE, pyproject.toml

Who it is for

  • Stated purpose · manifest declares · unreviewed

    🤗 smolagents: a barebones library for agents. Agents write python code to call tools or orchestrate other agents.

    pyproject.toml:8

What it can do

No explicit capability or integration was extracted; that does not mean none exists.

What adoption takes

  • Stated prerequisite · manifest declares · unreviewed

    Manifest declares Python >=3.10 as a requirement.

    pyproject.toml:13
  • Declared dependency · manifest declares · unreviewed

    Declares a dependency on huggingface-hub.

    pyproject.toml:15
  • Declared dependency · manifest declares · unreviewed

    Declares a dependency on requests.

    pyproject.toml:16
  • Declared dependency · manifest declares · unreviewed

    Declares a dependency on rich.

    pyproject.toml:17
  • Declared dependency · manifest declares · unreviewed

    Declares a dependency on jinja2.

    pyproject.toml:18
  • Declared dependency · manifest declares · unreviewed

    Declares a dependency on pillow.

    pyproject.toml:19
  • Declared dependency · manifest declares · unreviewed

    Declares a dependency on python-dotenv.

    pyproject.toml:21

Limits and terms

  • Stated limitation · repository says · unreviewed

    Security is a critical consideration when working with code-executing agents. Ensure you are using one of the sandboxed execution options that provide isolation from untrusted code.

    README.md:270
  • Declared license · manifest declares · unreviewed

    A license or copying file is present; its terms have not been reviewed.

    LICENSE:1

Observed technical context

Python · 48 selected filesAgent orchestration reference · static onlyModel API reference · static onlyTool-calling reference · static only

These are references in selected source files, not proof that a feature works at runtime.

Observed in the scanned archive at Oct 9, 2026, 2:19 PM UTC. Archive digest sha256:73523d8aaa4dd593…. Source links open current GitHub HEAD and may differ from the scan. Repository statements are unreviewed and do not affect the score.

02 / FINDINGS FIRST

What needs
a closer look.

Critical details are withheld publicly for the first 30 days after a scan. A matched pattern is a lead for investigation, not a confirmed vulnerability.

7 findings
CRITICAL

Critical finding · details withheld for 30 days

SEC001 · Security & privacy

Details are withheld from public view for 30 days after the scan. The owner can see them in full.

SUGGESTED ACTION

Review the private assessment and address the finding.

[location embargoed] · location unavailable
CRITICAL

Critical finding · details withheld for 30 days

SEC001 · Security & privacy

Details are withheld from public view for 30 days after the scan. The owner can see them in full.

SUGGESTED ACTION

Review the private assessment and address the finding.

[location embargoed] · location unavailable
CRITICAL

Critical finding · details withheld for 30 days

SEC002 · Security & privacy

Details are withheld from public view for 30 days after the scan. The owner can see them in full.

SUGGESTED ACTION

Review the private assessment and address the finding.

[location embargoed] · location unavailable
MAJOR

LLM request without token ceiling

REL001 · Guardrails & loop control

A model request does not set a maximum output length. One unusual input can produce a very long, expensive response.

SUGGESTED ACTION

Set max_tokens or the provider's equivalent on each model call.

open_deep_research/scripts/mdconvert.py:756
MAJOR

LLM output without runtime schema

SCH001 · Schema & tooling

Model output is used without being checked against a schema. Unexpected or malformed output can break later steps or be trusted when it should not be.

SUGGESTED ACTION

Validate model output with Zod, JSON Schema, Pydantic or provider structured output.

open_deep_research/scripts/mdconvert.py:756
MAJOR

Model call without fallback

EFF001 · Token & cost efficiency

A model call has no visible fallback. If the provider fails or is rate limited, the whole agent fails with it.

SUGGESTED ACTION

Configure a secondary model or explicit fallback path for provider failures.

open_deep_research/scripts/mdconvert.py:756
MAJOR

Domain ownership not verified

TRUST001 · Entity trust

Metal Mantra could not confirm that the owner controls the stated domain. The domain may belong to someone else.

SUGGESTED ACTION

Verify the domain using the Metal Mantra DNS TXT challenge.

huggingface.co · location unavailable
03 / EVIDENCE & LIMITS

The boundary
is part of the result.

What was scanned
48 of 48 eligible files · Complete selection.
Supported static source patterns and domain checks; no runtime execution.
What remains unknown
Unscanned files, external dependencies, production configuration, actual agent behavior and evidence beyond the configured checks.
Source location
Links open the current repository HEAD. Source may have changed since this snapshot.
Something looks wrong?
Tell us if a finding, scope or claim on this report needs a correction. Report a problem with this report
Entity trust · 15% of v0.2
Includes DNS ownership, HTTPS and policy evidence. Verification is domain control, not a security certificate.
DNS: unverifiedHTTPS: availablePrivacy: foundTerms: found
04 / THE METHOD BEHIND THE NUMBER

Open the
calculation.

Each pillar starts at 100. Rule deductions reduce its score; weights determine the overall contribution.

01

Security & privacy

3 findings · 25% weight
46/10011.5 pts
02

Guardrails & loop control

1 findings · 25% weight
88/10022.0 pts
03

Schema & tooling

1 findings · 20% weight
84/10016.8 pts
04

Token & cost efficiency

1 findings · 15% weight
86/10012.9 pts
05

Entity trust

1 findings · 15% weight
50/1007.5 pts
Inspect the current standard
USER REPORTS

Did smolagents work for people?

No user reports yet.

Each report is one signed-in GitHub account saying what happened. We do not verify that the account ran the agent, and reports never change the automated score or grade. For a security problem, use the disclosure page instead of a public note.

Sign in with GitHub to add a report

SHARE THE EVIDENCE

A signal.
Not a seal.

Share the full report so its scope and limits travel with the score.

Metal Mantra automated signal 71/100, not certified