Models · Inference.net
Pluely × Inference.net: Schematron V2 Turbo
Text → TextNo setup · no API key
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
How Inference.net: Schematron V2 Turbo works in Pluely
- One tap to switch. Open the overlay's model picker, search "Inference.net: Schematron V2 Turbo", click — your next question uses it. See the model picker.
- Works everywhere in the app. Ask mode questions, automatic responses in meetings, transcript summaries, and follow-ups all run on whichever model you selected.
- Text-first. Inference.net: Schematron V2 Turbo is a text model — Pluely automatically disables the screenshot controls while it's selected so you never send an image a model can't read.
- Metered simply. Each completion counts as one AI request on your monthly plan — see plans & usage.
- Private by design. Conversations stay on your device, and Pluely never exposes its serving infrastructure to the app.