FideAI

Research

Research for trustworthy AI.

Fide AI studies AI systems through the lens of making them more trustworthy. We research whether systems use evidence well, remain within legitimate authority, preserve human control, and stay accountable while they act.

This research develops and validates the methods independent assurance depends on. We publish evidence and public-good artifacts that others can inspect, challenge, and use in their own work.

Explore by domain

Find the setting behind the research.

Our domain hubs bring together Fide’s work, future interests, and selected research from the field.

Explore 7 connecting research themes →

88 public research calls translate the agenda into questions others can take forward.

Explore the public agenda and contribute to the work in Fide AI Research on GitHub.

Published research

Released studies, with their context intact.

Our released studies currently focus on faith-facing systems. Each page states the question, method, result, artifacts, and limits so readers can inspect the evidence directly.

Research · Faith & Religious Life

8,640 matched requests

When Not to Generate

Can AI quote Scripture exactly? We tested four ways of producing a quotation and traced where each one can fail.

Source fidelity

Research · Faith & Religious Life

4,800 controlled observations

Knowing When to Defer

Will AI check an authoritative Scripture source when a user asks it to rely on memory?

Source fidelity & authority

Research in development

See how the research takes shape.

Inspect the questions and methods behind our current study. Enterprise is an exploratory next direction; other proposals remain available for collaborators to develop.

Cybersecurity

Protocol development

When Security Evaluations Go Stale

Fide is developing a study of whether targeted retesting catches meaningful performance declines, or gives false reassurance. We begin with AI malware-report analysis, comparing smaller retests against complete benchmark reruns.

Protocol and evaluation tooling in development. No model-performance findings yet.

Explore the study brief

Updated

Explore our proposed enterprise starting point →
Other proposals for collaborators

Interpretation limits

A result describes behavior under named versions, prompts, conditions, rubrics, and evaluation procedures. It does not establish general safety, product approval, or deployment readiness beyond the tested scope. Faith-domain studies also do not confer theological or pastoral authority.