Legacy / unsupported — superseded by FinStruct

Status as of 2026-07-27: this repository is retained for provenance and download continuity. It is not an active Uraion Labs product, is not used by FinStruct, and is not supported for production or financial-document workflows.

Why it was superseded:

  • The prior card said the model was benchmarked on BFCL-v4 and IFEval, but this repository contains no scores, raw predictions, evaluation script, base-model comparison, or environment manifest. That statement is withdrawn as unverified.
  • The root GGUF files are documented in the original card as having malformed tensor shapes because they were converted directly from NF4-packed weights. They are not recommended for llama.cpp, Ollama, or LM Studio.
  • The Transformers NF4 path has not been independently reproduced in a clean, versioned evaluation and has no measured advantage over its Qwen base model.
  • 7,794 Hub downloads were recorded at audit time. Download count does not prove successful loading, adoption, model quality, or customer use.

No stronger capability claim should be inferred from the retained files. The full pre-audit card is preserved in LEGACY_CARD.md for historical transparency.

Current Uraion Labs work is FinStruct: auditable, local-first extraction of SEC filings with versioned schemas, evidence, abstention, raw predictions, and reproducible benchmarks. See uraionlabs.com.

Downloads last month
3,851
GGUF
Model size
2B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

2-bit

3-bit

4-bit

5-bit

6-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for UraionLabs/Uraion-Agent-Small

Finetuned
Qwen/Qwen3.5-2B
Quantized
(145)
this model

Collection including UraionLabs/Uraion-Agent-Small