Go EuropeAIdirectory

Evidence-backed European AI register

Pleias

Reviewed identity, roles, model listings, pricing and evidence for Pleias.

Dataset compiled 2026-08-27 · evidence over inference

model labresearch institute9 model listings0 token-price rows

Reviewed identity

Headquarters scope
EU headquarters
Headquarters record
Paris (Station F, 5 Parvis Alan Turing, 75013)
Legal entity
SAS
Ownership
founder-owned private (no disclosed VC round)
Founded
2024
Confidence
medium
Last verified
2026-08-22

Links

Official website

Open interactive profile

Evidence boundaries

Organisation listings identify a reviewed canonical record. Model listings show an association in the source register; they do not by themselves establish model creation, current API availability, hosting location or infrastructure ownership.

Model listings

  1. Baguettotroncatalogued with this organisation · origin not asserted
    Type
    text
    Parameters
    0.3B
    Context
    unknown
    Languages
    multilingual (French-centric)
    Licence
    unknown
    Release
    2025-2026
  2. Monadcatalogued with this organisation · origin not asserted
    Type
    text
    Parameters
    56.7M
    Context
    unknown
    Languages
    unknown
    Licence
    unknown
    Release
    2025-2026
  3. Pleias-RAG-350M / Pleias-RAG-1Bcatalogued with this organisation · origin not asserted
    Type
    text
    Parameters
    0.35B / 1B
    Context
    unknown
    Languages
    multilingual
    Licence
    unknown
    Release
    2025
  4. Pleias-SLM-RAGcatalogued with this organisation · origin not asserted
    Type
    text
    Parameters
    0.3B
    Context
    unknown
    Languages
    unknown
    Licence
    unknown
    Release
    2026
  5. Silloncatalogued with this organisation · origin not asserted
    Type
    text
    Parameters
    0.6B
    Context
    unknown
    Languages
    French
    Licence
    unknown
    Release
    2026
  6. CommonLinguacatalogued with this organisation · origin not asserted
    Type
    text
    Parameters
    unknown
    Context
    unknown
    Languages
    multilingual
    Licence
    unknown
    Release
    2026
  7. Celadoncatalogued with this organisation · origin not asserted
    Type
    text
    Parameters
    0.1B
    Context
    unknown
    Languages
    unknown
    Licence
    unknown
    Release
    2024
  8. OCRerrcrcatalogued with this organisation · origin not asserted
    Type
    text
    Parameters
    0.4B
    Context
    unknown
    Languages
    unknown
    Licence
    unknown
    Release
    2025
  9. ksante-colbert-smallcatalogued with this organisation · origin not asserted
    Type
    embedding
    Parameters
    33.4M
    Context
    unknown
    Languages
    French
    Licence
    unknown
    Release
    2025

Profile note

The 'clean data' lab: coordinates Common Corpus (~2 trillion tokens, the largest open multilingual public-domain/rights-cleared pre-training dataset, ICLR 2026 oral) and trains tiny specialised models entirely on it. 31 models and 62 datasets published. No pricing is public — this is the main gap in the record.

Record sources

  1. https://pleias.ai/
  2. https://huggingface.co/PleIAs
  3. https://huggingface.co/datasets/PleIAs/common_corpus
  4. https://arxiv.org/html/2506.01732v3
  5. https://tracxn.com/d/companies/pleias/__hMwegh0N9UlX2IOJGgxdUJORGIIeEBf2eZ6QkNFM19U