Manifund foxManifund
Home
Login
About
People
Categories
Newsletter
HomeAboutPeopleCategoriesLoginCreate
alexey-safonov-liminal avataralexey-safonov-liminal avatar
Alex Safonov

@alexey-safonov-liminal

Independent AI safety researcher building deterministic oversight, trace evaluation, and causal memory tools for agent systems.

https://github.com/safal207
$0total balance
$0charity balance
$0cash balance

$0 in pending offers

About Me

I work on AI safety infrastructure for agent systems: deterministic trace protocols, causal memory layers, adversarial evaluation, and failure localization. Before pivoting into AI safety, I spent 12+ years in fintech QA, where I worked on high-reliability testing, auditability, and production-grade failure analysis. My goal is to turn vague concerns about agent misalignment into concrete, reusable open-source artifacts: benchmarks, trace tooling, and oversight libraries.

Projects

Preventing Causal Overclaiming in AI-Assisted Biologypending grant agreement signature
Deterministic Oversight for Agent Traces — LTP + CML

Comments

Preventing Causal Overclaiming in AI-Assisted Biology
alexey-safonov-liminal avatar

Alex Safonov

1 day ago

Initial progress update: working prototype and scientific outreach

The project begins with a working open-source prototype rather than a purely prospective research plan.

We have implemented an initial closed reasoning loop:

source evidence → Kairos evaluation → LiminalDB persistence → RINSE reinterpretation → Kairos revalidation

The first case study examines claims derived from research on archaic introgression. The system preserved the supported association-level interpretation while preventing automatic escalation to adaptive causality where evidence for expression changes, cellular effects, organism-level phenotype, and fitness advantage remained absent.

Current technical outputs include:

  • versioned and machine-readable scientific claims;

  • explicit missing-evidence and causal-gap records;

  • immutable preservation of earlier interpretations;

  • deterministic replay and revalidation;

  • source and Git commit pinning;

  • CI-backed receipts and tamper checks;

  • explicit research-only, execution, deployment, and authority boundaries.

We have also requested external feedback from researchers working in computational genomics, population genetics, archaic introgression, and adaptive introgression.

The purpose of this outreach is not to claim that we have independently reproduced the underlying biology. It is to test whether the claim-lineage and causal-boundary model accurately represents how domain experts evaluate and revise scientific conclusions.

Public development:

  • RINSE reflection-loop implementation:
    https://github.com/safal207/rinse/pull/23

  • Kairos independent revalidation:
    https://github.com/safal207/Kairos-Gate-for-X-Cell/pull/62

The next milestone is to incorporate expert criticism, simplify the researcher-facing workflow, and publish the first reproducible biological casebook.

Collaboration Infrastructure for Safeguarding AI
alexey-safonov-liminal avatar

Alex Safonov

3 months ago

Super

Deterministic Oversight for Agent Traces — LTP + CML
alexey-safonov-liminal avatar

Alex Safonov

3 months ago

PythiaLabs has now shipped a runnable pre-execution gate demo:

- make demo

- real Web3 treasury engine

- 4 scenarios

- SHA-256 evidence verification

- counterfactual rejected → accepted

- landing page with runnable proof

This strengthens the implementation pathway for the LTP + CML research roadmap.

Deterministic Oversight for Agent Traces — LTP + CML
alexey-safonov-liminal avatar

Alex Safonov

4 months ago

90-day execution plan

Days 1–30: finalize trace schema, integrate one framework, define failure taxonomy.

Days 31–60: build initial adversarial corpus, implement CML graph logic, run first baseline comparisons.

Days 61–90: expand benchmark, publish evaluation write-up, release reusable library + documentation.

Deterministic Oversight for Agent Traces — LTP + CML
alexey-safonov-liminal avatar

Alex Safonov

4 months ago

Update:

I have active related applications under evaluation with LTFF, Open Philanthropy / Coefficient-affiliated funding, and NLNet, and I am currently clarifying fiscal sponsorship structure for an external SFF application. Current focus is tightening benchmark scope, deliverables, and release plan for LTP + CML.

Deterministic Oversight for Agent Traces — LTP + CML
alexey-safonov-liminal avatar

Alex Safonov

4 months ago

Update:

I have active related applications under evaluation with LTFF, Open Philanthropy / Coefficient-affiliated funding, and NLNet, and I am currently clarifying fiscal sponsorship structure for an external SFF application. Current focus is tightening benchmark scope, deliverables, and release plan for LTP + CML.

Deterministic Oversight for Agent Traces — LTP + CML
alexey-safonov-liminal avatar

Alex Safonov

4 months ago

Update: We are aligning fiscal sponsorship for an external SFF application. This project is gaining multi-fund interest.