Skip to content
View mertkayacs's full-sized avatar
:electron:
Eschatialabs for the future of SI ✨
:electron:
Eschatialabs for the future of SI ✨

Organizations

@WaterForTansania

Block or report mertkayacs

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
mertkayacs/README.md

Eschatia Labs: open models, papers and software from the edge of what we know

Mert Kaya

Senior AI researcher and AI architect. I build small AI models and tools for decisions whose confidence and evidence people can inspect. I founded Eschatia Labs, an international open-source research lab, and work as an AI Engineer at Enocta.

Visit my site for project demos and research. Hugging Face hosts the models and datasets.

Models and evaluation

  • JevAlt: small language models for choosing between options in English, Turkish and German, with a probability for each answer. Karar-4B is one of the best open Turkish decision models at 4B parameters: 96.8% accuracy on held-out Turkish decisions against Kev-4B's 87.1%. The tests come from JevAlt's own data pipeline. Try the models.
  • JevOss: test a decision model's accuracy, confidence and response to hidden instructions. Test methods.
  • Open benchmarks: jevalt-bench has 3,116 written situations in English, Turkish and German for decision models, and Tholos-Bench has 160 workspace tasks for small AI assistants. Both are free to download and run against your own model.
  • xdfdet: eight deepfake video detectors with heatmaps showing the face regions behind a prediction, from my MSc thesis. Study and demos.

Tools for AI coding and local AI

  • Ultra Mod: an all-in-one mod pack for Claude Code. Usage limits and context above the prompt, a guard with undo for rm -rf and git reset --hard, .env files kept away from the model, and a receipt under every answer. Install with npx ultramod. Site and demos.
  • Tholos: AI assistants that share tables, notes and tasks on your computer, with rules that allow, ask about or deny actions. Tholos-2B passed 137 of 160 Tholos-Bench scenarios against MiniCPM5-2B's 112 on a Kaggle T4 with llama.cpp and JSON schema decoding. Model and benchmark.
  • reevesagents: run AI coding tools side by side in a terminal workspace, or let one direct the others. Demo and docs.

Research

MSc Computer Science, TED University. BSc Computer Engineering, TOBB University of Economics and Technology.

  • M. Kaya, V. Adanova. Augmentation and Cutout in Deepfake Detection: A Comparative Study of Accuracy, Calibration, and Attention. UBMK 2026. Accepted, to appear in IEEE Xplore.
  • M. Kaya. Explainable deepfake detection using frame level CNN models: A comparative study of augmentation and cutout techniques. MSc thesis, TED University, 2025. Advisor: V. Adanova. Read the thesis.
  • M. Kaya. Language Preservation Using Agentic AI Architectures: A Tool-Grounded Small Language Model on a Constructed-Language Testbed. Preprint, Eschatia Labs, 2026. Read the paper.

LinkedIn and Kaggle.

Pinned Loading

  1. reevesagents reevesagents Public

    Run AI coding tools side by side in a local tmux workspace. CLI, terminal UI, Web UI and MCP control for Claude Code, Codex, Kimi and more.

    TypeScript 88 8

  2. xdfdet xdfdet Public

    Explainable deepfake video detection with eight EfficientNet-B4 models, Grad-CAM heatmaps and facial-region analysis. Code from an MSc thesis and UBMK 2026 study.

    Python 8

  3. jevalt jevalt Public

    Three small 4B language models for local decisions in English, Turkish and German. Jev-compatible API, per-option probabilities and CPU GGUF builds.

    Python 5 1

  4. jevoss jevoss Public

    Evaluate Jev-compatible decision models: accuracy, probability calibration, prompt injection, option-order sensitivity and consistency.

    Python 2

  5. ultramod ultramod Public

    The best all-in-one mod pack for Claude Code: usage limits and context HUD, a guard with undo for rm -rf and git reset --hard, .env and secret protection, and a receipt for every turn. Ten mods, on…

    TypeScript 2

  6. tholos tholos Public

    Local AI assistants sharing tables, notes and tasks, with scheduled runs and allow, ask or deny rules. Includes the Tholos-2B tool-use model.

    Python 1 1