Skip to main content

attestq

Answer security questionnaires and compliance attestations from your own evidence.

attestq is a small, model-agnostic RAG kernel for a problem every security and GRC team has: you have a questionnaire (a vendor security review, a SIG/CAIQ response, an audit-evidence request, a due-diligence form, the security section of an RFP) and you have a pile of evidence (SOC 2 reports, policies, standards, prior questionnaires). attestq retrieves the relevant evidence for each question and drafts a grounded, cited answer — and, just as importantly, tells you plainly when the evidence isn't there.

It is not a classic ML / training problem. It's retrieval + an LLM you bring yourself. attestq owns the orchestration; you own the model, the embedder, and the store.

Why it exists

The same pattern keeps getting rebuilt one-off inside every program: chunk the docs, embed them, retrieve per question, prompt an LLM, paste into a form. Done naively it hallucinates (answers with no supporting evidence) and it silently drops the one focused document that actually had the answer. attestq bakes in the two hard-won fixes:

  • Confidence gate. When the best retrieved evidence scores below a threshold, the question is answered "insufficient evidence" without calling the LLM. Absence of evidence is a first-class, valid result — not an invitation to guess.
  • Wide rerank window. The kernel keeps a generous number of chunks after reranking so a single relevant document isn't dropped on a small corpus.

Install

pip install attestq                 # dependency-free core
pip install "attestq[chroma]"       # + persistent Chroma vector store
pip install "attestq[openai]"       # + OpenAI-compatible chat adapter
pip install "attestq[ollama]"       # + local Ollama embeddings
pip install "attestq[rerank]"       # + cross-encoder reranker
pip install "attestq[loaders]"      # + pdf / docx / xlsx loaders
pip install "attestq[all]"          # everything

The core has zero third-party dependencies. Adapters are opt-in extras.

Quick start

You inject any chat model and any embedder as plain callables:

from attestq import Engine, Question

# Bring your own model + embedder (one-liners around any provider).
def my_chat(prompt: str) -> str:
    ...   # call OpenAI, Anthropic, a local model, your corporate gateway, ...

def my_embed(texts):
    ...   # return one vector per text

engine = Engine(chat=my_chat, embed=my_embed)   # in-memory store by default

# Ingest a vendor's evidence into its own namespace.
engine.ingest(
    [
        ("All customer data at rest is encrypted with AES-256...", {"source": "DataProtection.pdf"}),
        ("MFA is enforced for all privileged access...", {"source": "AccessControl.docx"}),
    ],
    namespace="helios",
)

# Answer one control.
ans = engine.evaluate(
    Question(
        id="ENC-1",
        prompt="Is customer data encrypted at rest?",
        choices=["Met", "Not Met", "Not Applicable"],
    ),
    namespace="helios",
)

print(ans.determination)          # "Met"
print(ans.confidence)             # 0.0 - 1.0 retrieval confidence
print(ans.insufficient_evidence)  # False
for c in ans.citations:
    print(c.source, "->", c.snippet)

Run a whole questionnaire:

from attestq import Questionnaire

qn = Questionnaire(
    id="vendor-sec-review",
    title="Vendor Security Review",
    questions=[
        Question(id="ENC-1", prompt="Is data encrypted at rest?", choices=["Met", "Not Met", "Not Applicable"]),
        Question(id="IAM-1", prompt="Is MFA enforced for privileged access?", choices=["Met", "Not Met", "Not Applicable"]),
    ],
)

answers = engine.evaluate_all(qn, namespace="helios", on_answer=lambda a: print(a.question_id, a.determination))

Command line

Installing attestq gives you an attestq command. Run the bundled sample, or point it at your own questionnaire and evidence:

# Run the bundled fictional sample assessment end-to-end
attestq demo -o report.md

# Evaluate your questionnaire against a folder of evidence
attestq run -q questionnaire.yaml -e ./vendor-evidence -n acme -o report.docx

# Pipe JSON to another tool
attestq run -q q.json -e ./evidence --format json | jq .summary

Providers are resolved from flags or environment, so the same command works against OpenAI-compatible endpoints or a local Ollama:

export OPENAI_API_KEY=sk-...                 # uses OpenAI by default
attestq demo

attestq demo --provider ollama               # local, no key, nothing leaves the host
attestq run -q q.yaml -e ./ev --provider openai --base-url https://my-gateway/v1

Try it instantly (no setup)

The built-in HashEmbedder needs no model and no service, so you can watch the retrieval pipeline and the confidence gate work the moment you install:

pip install attestq
python examples/quickstart.py

See examples/ for a full provider-wired demo (helios_demo.py) and a minimal web UI (examples/web/) that runs the sample assessment in your browser.

How it works

evidence docs ──▶ chunk ──▶ embed ──▶ vector store (namespaced per corpus)

question ──▶ embed ──▶ retrieve(k) ──▶ rerank(top_k) ──▶ confidence gate
                                                              │
                              below threshold ───────────────┤──▶ "insufficient evidence" (no LLM call)
                                                              │
                              above threshold ──▶ prompt LLM ─┴──▶ parse ──▶ Answer(determination, summary, citations, confidence)

Verifying the draft

Drafting is half the job; the other half is not trusting the draft. Turn on verify=True and every answer carries two reports:

engine = Engine(chat=my_llm, embed=my_embedder, verify=True)
answer = engine.evaluate(question, namespace="vendor-x")

if answer.needs_review:
    print(answer.grounding.unverified)   # ['2019-03-01'] — asserted, not in evidence
    print(answer.quality.detail())       # "only 20% of the answer is carried by ..."

answer.grounding extracts every concrete value the draft asserted — dates, versions, percentages, standards, durations — and confirms each one actually occurs in the evidence. This catches the highest-liability failure: "certified to ISO 27001 since 2019-03-01" when neither appears in a single retrieved document.

answer.quality catches the quieter failure — an answer that restates the question and cites evidence it never drew on:

Q: Are encryption keys securely managed through defined key-management processes? A: Yes. Encryption keys are securely managed through defined key-management processes.

That asserts nothing the question didn't already say. Per sentence, it measures support (closeness to the retrieved evidence) against echo (closeness to the question), each judged against the question's own similarity to that evidence — so there is no magic cosine constant to recalibrate when you swap embedders.

Both checks are deterministic: no second LLM grading the first, which would only share the drafter's blind spots. Both flag and never rewrite — a flagged answer may still be the right answer, so the disposition stays with the reviewer. Verification is off by default; it costs one extra embedding call per answer.

One judgment measurement genuinely can't make is whether a question wanted documentary evidence at all — "What is your registered legal entity name?" has a correct answer no policy document will ever back. Inject a small model to decide, and it may only ever clear a flag, never raise one:

from attestq import make_claim_classifier

engine = Engine(chat=big_model, embed=embedder, verify=True,
                claim_classifier=make_claim_classifier(small_fast_model))

Every failure path in that classifier — endpoint down, unparseable reply — leaves the flag standing.

Closing the loop

Every reviewer who edits a draft before shipping it is handing you a quality label for free — and most systems throw it away, because the final answer overwrites the draft in place. Keep both and you can answer the question every RAG stack should be able to answer and usually can't: does the confidence number mean anything?

from attestq import outcome_from_answer, build_scorecard

outcomes = [
    outcome_from_answer(answer, question, final=what_the_reviewer_shipped)
    for answer, question, what_the_reviewer_shipped in reviewed_items
]

card = build_scorecard(outcomes)
card["acceptance_rate"]   # 0.62 — shipped as drafted
card["calibrated"]        # True / False / None ("not enough data to say")
card["source_trust"]      # worst-performing documents first

Three things come out of it:

  • Acceptance rate — how often a draft ships as written, with edits graded from cosmetic to wholesale rewrite. Items still awaiting review count as neither, so the number doesn't move with queue depth.
  • Calibration by confidence band — the real test of your gate. A healthy profile is monotonic: acceptance climbing with confidence. Flat means the score is decorative and every threshold built on it is arbitrary; inverted means it's actively misleading. is_calibrated returns None rather than a verdict when the sample is too thin to support one.
  • Per-source trust — which documents back answers that survive review and which back rewrites. A document that repeatedly produces rewritten answers is stale, ambiguous, or wrong for the questions it keeps matching. This is the raw material for source-authority weighting at retrieval time.

outcome_from_answer is duck-typed, so records from a system that never touched attestq fold into the same scorecard as long as they expose the same handful of attributes.

Everything is swappable:

Piece Default Swap for
Chat model you inject it any LLM / gateway
Embedder you inject it Ollama, sentence-transformers, OpenAI
Vector store InMemoryVectorStore attestq[chroma]
Reranker none attestq[rerank] cross-encoder
Prompt / parser evidence-only default your own prompt_builder / response_parser

Design principles

  • Bring your own model. No provider is hard-wired. A lambda is enough.
  • Grounded or silent. Answers cite their evidence; thin evidence yields an explicit "insufficient" result, never a confident guess.
  • The model proposes, deterministic code disposes. Verification is plain string and vector math, so the verifier cannot hallucinate the way an LLM-grading-an-LLM judge can. It flags for a human; it never rewrites.
  • Per-corpus isolation. One store, many namespaces — keep each vendor's evidence separate without standing up a new index each time.
  • Light core. The kernel imports nothing third-party; heavy deps stay in extras you opt into.

Status

Usable today: the core kernel, the verification and feedback layers, in-memory + Chroma stores, OpenAI/Ollama adapters, a cross-encoder reranker, document loaders, JSON/Markdown/Word export, a CLI, and a web demo. Contributions and issues welcome.

License

MIT

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

attestq-0.4.0.tar.gz (73.5 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

attestq-0.4.0-py3-none-any.whl (59.3 kB view details)

Uploaded Python 3

File details

Details for the file attestq-0.4.0.tar.gz.

File metadata

  • Download URL: attestq-0.4.0.tar.gz
  • Upload date:
  • Size: 73.5 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/7.0.0 CPython/3.13.14

File hashes

Hashes for attestq-0.4.0.tar.gz
Algorithm Hash digest
SHA256 ad71aa9ac6a71f4c65a66da9a9fbd590775260fe2c0020010e8e810ef5533d07
MD5 f04a90a34b439d66689caf4956935549
BLAKE2b-256 355b3ed5e5ac24fbf7dec25efd1ff766b3edc5e2f196d55a5e338c971fef05e3

See more details on using hashes here.

Provenance

The following attestation bundles were made for attestq-0.4.0.tar.gz:

Publisher: publish.yml on vinayvobbili/attestq

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file attestq-0.4.0-py3-none-any.whl.

File metadata

  • Download URL: attestq-0.4.0-py3-none-any.whl
  • Upload date:
  • Size: 59.3 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/7.0.0 CPython/3.13.14

File hashes

Hashes for attestq-0.4.0-py3-none-any.whl
Algorithm Hash digest
SHA256 7246e6ce096c1493cdc4a8087ab13a0eb2838f1a0e840dce26e8a32b0a4d52ba
MD5 cd554b4ed05ef70015d0e52783d7a0cb
BLAKE2b-256 93f6970401a0d61d2be323fdc77cdb42ea7ec1b79fc860210bb2e909b47775cf

See more details on using hashes here.

Provenance

The following attestation bundles were made for attestq-0.4.0-py3-none-any.whl:

Publisher: publish.yml on vinayvobbili/attestq

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page