
Guide
What Early Open Weight Model Deployments Reveal About Production AI
Companies are deploying open weight models in production. We examine early deployment strategies, evaluation practices, and system design.

Guide
The Best Open Weight Models of Summer 2026
A recap of the best open-weight models of summer 2026 and which model to reach for depending on the task.

Research
How GLiNER Grew From NER Into Structured Extraction
From a PhD project to 45 million downloads: how GLiNER became GLiNER2.5, and why encoders still beat LLMs at extraction.

Research
Introducing GLiNER2.5: Efficient Span-Free Information Extraction with Schema-Driven Interface
A new boundary-prediction architecture replaces span enumeration, adding joint entity-relation extraction, constrained classification, and unlimited span length.

Guide
How to Use Open-Weight Models in 2026: A Developer's Guide
Open-weight models now rival the frontier, closed models. Compare Kimi K3, DeepSeek V4, GLM 5.3, Qwen 3.8 and more, with pricing, benchmarks, and how to switch.

Research
Introducing Fastino-Nemotron-3.5-Lightning-Finance and Fastino-Nemotron-3.5-Lightning-Healthcare
Two specialized open weight models for regulated industries, fine-tuned on NVIDIA Nemotron 3.5 Lightning by the Fastino Fine-Tuning Agent.

Research
Small Model, Big Leverage: What We Learned Fine-Tuning NVIDIA Nemotron 3.5 Lightning with an Autonomous Agent
Learnings from fine-tuning Nemotron 3.5 Lightning with an autonomous agent.

Guide
How to fine-tune open weights models: Model selection, data curation, fine-tuning strategies, and evals
An overview on how to fine-tune open-weight models, popular model options, and fine-tuning strategies.

Research
Loop Engineering Needs a Smarter Inference Layer
What is loop engineering, how it breaks without automated model routing, and why routing needs to live in the inference layer.
Read our published research, written by our technical team, from new model releases to product research.
Read the papers

Subscribe to our newsletter
Be the first to hear about what's new at Fastino Labs, including new model releases, product launches, and upcoming events.
Fastino Inc. (“Fastino”) develops specialized AI models and provides APIs designed to support structured data extraction, classification, reasoning, and production AI workflows. Fastino is a technology company and does not provide legal, financial, compliance, or advisory services.
Any outputs, predictions, classifications, or decisions generated through Fastino models are based on the configuration, data, and implementation provided by the customer. Fastino does not control, verify, or guarantee the accuracy, completeness, or suitability of model outputs for any specific purpose. By using this website or Fastino’s models and services, you acknowledge that all content and outputs are provided for informational and operational purposes only and agree to our Terms of Use and Privacy Policy.
2026 Fastino Inc.
All rights reserved