ShinrAI
INNOVIUS.AI
Open beta — now enrolling

ShinrAI Encryption Models

Semantic encryption for the AI age: open-weight models that shield who you are before any AI sees your text, with keys that stay yours. Developed with the EECC Research Labs, trained at the Jülich Supercomputing Centre — released for everyone.
You use the AI. The AI does not use you.

🛡️ Semantic encryption, multi-dimensional 🔓 Apache 2.0 open weights 🌍 DE · EN · JA available today 💻 Runs on your hardware 🖥️ Trained on JURECA & JUPITER
The Models

Semantic Encryption, Not Redaction

Compact encoder models (ModernBERT family) — the detection layer of ShinrAI semantic encryption. Privacy protection that works in multiple dimensions, across languages and document styles.

DE Robert wohnt seit 2019 in Köln. Paul wohnt seit 2019 in Hamburg.
EN Emma Wilson moved to Birmingham. Sophie Turner moved to Glasgow.
JA 山田さんは大阪に住んでいます。 佐藤さんは名古屋に住んでいます。

No black bars. Replacements keep culture, rarity and gender intact — so not even an AI can tell.

🎭 Multi-dimensional replacement

Classical tools redact ([REDACTED] wreckage) or hard-cut. ShinrAI substitutes are matched to context, culture and language at once — the AI reasons on, identities stay hidden.

⚖️ Bias-safe by design

Every substitute is statistically weighted so it never accidentally triggers biases or red flags downstream.

🔑 Keys stay yours

The matching key pair never leaves your device, your backend, or your instance. Only the key holder can map protected text back to reality.

🧅 Onion-routed access

The ShinrAI software optionally onion-routes requests, separating who asks from what is asked — across OpenRouter and any other access path. More paths, stronger protection.

🏥 Protection at scale

De-identify electronic patient files (ePA) for pharma research, clinical archives, high-security environments and banking — data stays analytically useful.

🔓 Open weights & data

Apache 2.0 weights; license-clean training-data releases (the fully synthetic corpus tracks license provenance per record — no proprietary APIs anywhere).

ChtSafe

Need this protection now? A compact DE/EN/JA edition already runs in production — ChtSafe for individuals, the Secure AI Suite for organizations (on-premise first, SaaS-hosted on request).

Talk to us

● Available today

GermanEnglishJapanese

A compact edition of these models already protects production traffic in our commercial products — ChtSafe for individuals, the Secure AI Suite for organizations (on-premise first, SaaS on request).

◐ In training — the open-weight suite

ItalianFrenchSpanishPolishPortuguese (PT/BR)RussianUkrainianTurkishKoreanArabicDutchSwedishChinese (Simpl.)Your language?

Languages land as they pass our quality gates — the form below directly shapes the order. Tell us what you need.

The Beta

What You Get — and What We Ask

We slot a limited beta cohort by language, hardware and domain, so every configuration gets real coverage.

Relaxed private AI use at home — protected by ShinrAI

Everyday AI, without giving yourself away.

⬇️ Early checkpoints

Model checkpoints and quantized builds before the public release, matched to the hardware you tell us about.

🧭 Real influence

Your language wishes, domains and votes feed directly into training priorities. This is genuinely how we decide what to build next.

💬 A direct line

Found a miss, a false positive, an awkward replacement? Beta feedback goes straight to the team training the next generation.

🎁 Open at the end

Everything lands as open weights right here on Hugging Face — beta participants simply get there earlier and shape what "there" looks like.

Join

Apply for the Beta — or Just Stay in the Loop

Two minutes helps us slot you into the best-fitting cohort. Or leave just an email and we will ping you once at release.

We will send you a single email when the ShinrAI encryption models are publicly released — no newsletter, no follow-ups. You can also follow Innovius on Hugging Face to see every release in your feed.

About you

Languages

Which languages does your AI usage mostly happen in? Pick everything that applies — this decides which language models we hand you first.

Your hardware

Where would you run the models? This drives which builds and quantizations we prepare for you.

Quantization

Interested in specific quantized builds? Pick any — or let us recommend.

Where would you use them?

Areas of use — pick everything that applies.

Vote: where should we specialize first?

How to read this: the Common models already cover everyday text — and with it most other areas — to a solid degree. Specialization matters where the language itself changes: government and aerospace documents carry their own vocabulary and identifier styles, and for coding a specialist is essential — personal data hides in source code, logs and configs in ways prose models don't see. Our roadmap so far: Common ships first, Space Exploration & Aerospace Engineering is a promise we will keep regardless of votes (we love it too much not to), and NSFW comes last. Your vote reorders the middle.

Pick up to 3 areas, in the order you care about them.

Coding it is! Which languages should the coding specialist understand first?

Anything else?

Application received — thank you!

We will review your application and get back to you as we slot the next beta cohort. Meanwhile: follow Innovius on Hugging Face and @InnoviusAI — or talk to us if you need ShinrAI protection in production today.

Credits

Built With — and Thanks To

This programme exists because remarkable institutions and open projects make serious research possible outside big tech.

EECC — European EPC Competence Center

EECC Research Labs

The ShinrAI encryption models are developed by Innovius together with the research lab of the European EPC Competence Center (EECC) — long-time partners in applied AI and privacy research.

eecc.info · EECC on Hugging Face

Forschungszentrum Jülich logo

Forschungszentrum Jülich — JSC

Training runs on the JURECA supercomputer, with the programme scaling onto JUPITER — Europe's first exascale system — at the Jülich Supercomputing Centre, supported through the WestAI initiative. Our deepest thanks to FZ Jülich and the JSC team: this work is only possible because Europe's research infrastructure is open to projects like ours.

fz-juelich.de/jsc · westai.de

Open models we build on

The base encoder is mmBERT (Johns Hopkins CLSP, MIT) from the ModernBERT family. Training data is generated, cross-checked and evaluated by openly released models from the Qwen (Alibaba), Gemma (Google DeepMind), Mistral and NVIDIA Nemotron families — thank you for keeping frontier-quality open models available; this project runs no proprietary APIs at all.

mmBERT · MITModernBERTQwenGemmaMistralNemotron

Open data

Ground truth is anchored in open data: GeoNames (CC-BY), Wikidata (CC0), the US SSA & Census name statistics (public domain), and national open-data sources including INSEE (France), INE (Spain), the Polish PESEL registry statistics, and Italian & German municipal open data. Open data is what makes honest, bias-aware ground truth possible.

GeoNames · CC-BYWikidata · CC0SSA / Census · PDINSEE · INE · PESEL

The open-weights pledge

Open weights are a conviction, not a marketing angle. We fully support the Open Weights and American AI Leadership open letter published in July 2026 by an NVIDIA-led coalition of 50+ organizations: models whose weights anyone can download, inspect and run on their own infrastructure are defensive assets — for security, for competition, for trust. The ShinrAI encryption models are our contribution from Europe: open, inspectable privacy infrastructure you can run yourself.