Research
Everything we publish, grouped by area.
Deepfake detection
Detectors that show where they look and how far their scores can be trusted.
Publications
-
An EfficientNet-B4 deepfake detector trained under nine augmentation and cutout configurations on FaceForensics++, scored on AUC, F1, Brier score and log loss, with Grad-CAM attention over facial regions. The four metrics do not agree on one best configuration.
Accepted at UBMK 2026. The final version will appear in IEEE Xplore.
Projects
-
Code and eight trained EfficientNet-B4 deepfake detectors from the thesis, with Grad-CAM region analysis and a public site with interactive detection and estimation games.
Code: MIT. Weights: CC BY-NC 4.0, under FaceForensics++ terms.
Decision models
Small models that answer questions in English, Turkish and German with calibrated probabilities, on a laptop CPU.
Publications
Projects
-
Three open decision models for English, Turkish and German. Each answer comes with a calibrated probability, and the models run in about 3 GB of RAM.
-
A medieval village in the browser where the JevAlt models decide how villagers respond to storms, fires, markets and wolves at dusk.
Models
Datasets
-
Training data for the JevAlt decision models.
-
Recorded scenes and films from Emberwick.
Local AI and open tools
Agents and tools that run on your own machine.
Publications
Projects
-
Small agents sharing a local workspace of tables, notes and a task board, with allow, ask and deny rules for each action.
Models
Datasets
-
Executed agent trajectories used to train Tholos-2B.
Language preservation
One of the best research efforts on language preservation with small local models.
Projects
-
A small language model that reads and writes Quenya and Sindarin, including Tengwar script, grounded in a grammar-verified corpus, with tools for translation, lookup, inflection and Tengwar rendering.
Code and weights are private. Demo by invitation.
Public benchmarks
Open tests and their results.
-
An open test bench for decision models using the Jev API, measuring accuracy, probability calibration and failure cases.
-
Benchmark and live-test results for JevAlt and the comparison models.
-
160 scenarios for local agents, each scored on the final state of the workspace.