Research & benchmarks

Research for AI writing tools that should help people learn.

A growing collection of transparent, reference-based studies on local AI writing models, multilingual feedback and real Windows performance. The goal is to provide practical evidence about the strengths, trade-offs and differences observed between models in controlled writing-coach scenarios.

Latest study

Qwen3 writing coach benchmark: 4B vs 8B vs 14B

Three local Qwen3 models received the same 20 writing-coach cases: correct English text, identify the real errors and explain the changes in French. The study also measured cold-start execution with Ollama on Windows.

Exploratory local benchmark

Which Qwen3 size offers the best practical balance for writing feedback?

See complete-case correction, expected-correction coverage, error localization, explanation-language compliance and cold-start performance—with the limits shown beside the results.

3 models20 paired cases60 completed responsesOllama on Windows

Research principles

Practical evidence for informed local-model choices.

Reference based

Models are compared against a frozen correction reference

The evaluated models do not see the expected corrections. Their outputs are compared offline with the same versioned cases and acceptance rules.

Multidimensional

Correction, explanation and execution are separated

A fast model is not automatically a better coach. Correction success, language compliance, structured output and local speed are shown as different dimensions.

Paired cases

The same cases are used for every model

Pairing helps reveal whether a visible score difference comes from many cases or from one isolated example.

Limits visible

Results are interpreted within the tested setup

Each study identifies its dataset, language pair, hardware and execution conditions so readers can understand where the findings are most useful.

From research to daily writing

Use the local model that fits your own Windows setup.

LinguaPilot lets you select text in almost any Windows app, receive a corrected version, an improved version and an explanation, and choose Ollama for local processing or an optional cloud provider for other tasks.