Blog

Blog

Latest updates

Guide

What Early Open Weight Model Deployments Reveal About Production AI

Companies are deploying open weight models in production. We examine early deployment strategies, evaluation practices, and system design.

Guide

The Best Open Weight Models of Summer 2026

A recap of the best open-weight models of summer 2026 and which model to reach for depending on the task.

Research

How GLiNER Grew From NER Into Structured Extraction

From a PhD project to 45 million downloads: how GLiNER became GLiNER2.5, and why encoders still beat LLMs at extraction.

Research

Introducing GLiNER2.5: Efficient Span-Free Information Extraction with Schema-Driven Interface

A new boundary-prediction architecture replaces span enumeration, adding joint entity-relation extraction, constrained classification, and unlimited span length.

Guide

How to Use Open-Weight Models in 2026: A Developer's Guide

Open-weight models now rival the frontier, closed models. Compare Kimi K3, DeepSeek V4, GLM 5.3, Qwen 3.8 and more, with pricing, benchmarks, and how to switch.

Research

Introducing Fastino-Nemotron-3.5-Lightning-Finance and Fastino-Nemotron-3.5-Lightning-Healthcare

Two specialized open weight models for regulated industries, fine-tuned on NVIDIA Nemotron 3.5 Lightning by the Fastino Fine-Tuning Agent.

Research

Small Model, Big Leverage: What We Learned Fine-Tuning NVIDIA Nemotron 3.5 Lightning with an Autonomous Agent

Learnings from fine-tuning Nemotron 3.5 Lightning with an autonomous agent.

Guide

How to fine-tune open weights models: Model selection, data curation, fine-tuning strategies, and evals

An overview on how to fine-tune open-weight models, popular model options, and fine-tuning strategies.

Research

Loop Engineering Needs a Smarter Inference Layer

What is loop engineering, how it breaks without automated model routing, and why routing needs to live in the inference layer.

Research Papers

Research Papers

Read our published research, written by our technical team, from new model releases to product research.

Read the papers

Subscribe to our newsletter

Be the first to hear about what's new at Fastino Labs, including new model releases, product launches, and upcoming events.

Fastino Inc. (“Fastino”) develops specialized AI models and provides APIs designed to support structured data extraction, classification, reasoning, and production AI workflows. Fastino is a technology company and does not provide legal, financial, compliance, or advisory services.

Any outputs, predictions, classifications, or decisions generated through Fastino models are based on the configuration, data, and implementation provided by the customer. Fastino does not control, verify, or guarantee the accuracy, completeness, or suitability of model outputs for any specific purpose. By using this website or Fastino’s models and services, you acknowledge that all content and outputs are provided for informational and operational purposes only and agree to our Terms of Use and Privacy Policy.

2026 Fastino Inc.

All rights reserved