Skip to content
View aalvsz's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report aalvsz

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
aalvsz/README.md

Ander Alvarez Sanz

AI systems engineer focused on verifiable LLM inference, computer-vision deployment, model compression, and agent reliability. I work as an AI Engineer at Multiverse Computing, turning research ideas into measured, reproducible systems for real hardware.

Completed work

  • ProvenanceGuard — source-aware factuality verification for MCP-based LLM agents, published on arXiv.
  • LibreYOLO — open-source object-detection library; I contributed the merged Core ML export implementation.
  • Agent Knowledge Vault — local-first, redacted knowledge capture for preserving Codex CLI and Claude Code session knowledge in Obsidian.
  • TurboQuant CPU — CPU-focused TurboQuant implementation work with modified llama.cpp builds and reproducible benchmark artifacts for x86-64 Linux and Apple silicon.

Ongoing work

  • Vision BuildKit — provider-neutral tooling for building, validating, and packaging computer-vision models. Its evidence records distinguish verified deployments from failed or incomplete backend gates across ONNX Runtime, OpenVINO, TensorRT, and Axelera.
  • Fused Memory — scoped multi-signal retrieval for long-term agents, evaluated through 61,600 paired retrieval cases with reproducible benchmark artifacts.
  • iOS calorie and exercise tracker (in development) — currently designing a native iOS app for calorie intake and exercise tracking.
  • DSpark Speed Lab (most recent project) — currently improving DSpark inference speed through exactness-first speculative-decoding research with frozen baselines, offline replay, matched runtime gates, and reproducible experiment protocols.

Publications

Focus

  • LLM inference, speculative decoding, and agent systems
  • Computer-vision deployment and hardware-aware validation
  • Model compression, quantization, and on-device AI
  • Reproducible benchmarks and evidence-gated engineering

Tools

Python, PyTorch, Swift, C++, ONNX Runtime, OpenVINO, TensorRT, Core ML, llama.cpp, and ML experiment tooling.

Contact and research

Pinned Loading

  1. vision-buildkit vision-buildkit Public

    Provider-neutral build, semantic conformance, and reproducible packaging for computer vision models

    Python

  2. fused-memory-harness fused-memory-harness Public

    Multi-signal retrieval and reproducible benchmark for long-term agent memory

    Python

  3. provenanceguard provenanceguard Public

    Source-aware factuality verification for MCP-based LLM agents — paper artifact

  4. agent-knowledge-vault agent-knowledge-vault Public

    Local-first Codex CLI and Claude Code skill for durable, redacted Obsidian knowledge

    Python

  5. dspark-speed-lab dspark-speed-lab Public

    Exactness-first fast iteration lab for improving DSpark inference speed

    Python

  6. LibreYOLO/libreyolo LibreYOLO/libreyolo Public

    LibreYOLO is a MIT licensed open source computer vision library

    Python 569 40