Scattered files
Important documents split across drives, email inboxes and local folders, with no single point of access and no search that finds them all.
Theka is AI document management for companies with a lot of documents: contracts, datasheets, manuals, certificates, product photos. It reads them all, tables and images included, and answers you citing document and page.
Stainless steel centrifugal pump for food processing plants. Operates at up to 6 bar1 and must be serviced every 12 months2.
Theka is Italian document management software for companies (DAM) with artificial intelligence. It gathers documents in a versioned archive with permissions and approvals, reads them (text, tables, images, scans) and makes them searchable by meaning. On recurring topics it generates wikis with verifiable citations. It works with your storage (S3, SharePoint, MinIO, SFTP) and with the AI model you prefer.
In many companies documents are scattered across shared folders, email, SharePoint, the ERP and personal computers. Finding the latest version of a contract, knowing which certificate is still valid or retrieving a figure from a hundred-page manual takes hours. And newcomers spend weeks working out where to look and which document is authoritative.
Important documents split across drives, email inboxes and local folders, with no single point of access and no search that finds them all.
Drafts, revisions and signed copies sit side by side: nobody is sure which is the current version, or who approved it.
Who read, edited, downloaded or shared a document? Without a reliable log, compliance becomes a risk.
An orderly archive, knowledge that builds itself, complete control over who sees what.
Every document with its versions, its status and its history.
Theka finds the topics that keep coming up and writes their wiki.
Every search and every answer respects the permissions of whoever is asking.
Drag your files in, connect SharePoint, S3 or an SFTP folder, or send them via API from your ERP. Even thousands of documents at once, with a time and cost estimate before you start.
Text, tables, charts and photos, with OCR where needed. It fills in the fields, writes the summary, finds the topics and flags when content and title don't match.
Search in your own words, not by file name. The answer comes with document and page, and only from the documents you are allowed to open.
On one side, the traditional company archive made of folders and email; on the other, document management software that knows what is inside the documents.
Theka combines seven signals (words, meaning, customers, projects, topics) and shows you the right passage, with the page.
Continuous updates, designed with the people who use Theka every day. The details are in the release notes.
Send a document for signature to several people, in order, and the signed PDF comes back to the archive on its own.
For each category, choose whether its facts join the knowledge of the whole company or stay with those who can open the documents.
If a document's content contradicts its title, type or category, Theka flags it: you correct it or dismiss it as a false alarm.
Thousands of files in batches: an estimate before you start, pause, resume and retry only the ones that failed.
Theka reads the data inside tables and charts and describes photos, so that information can be found and cited too.
Your product catalogue or ERP reads and writes in Theka with a dedicated key and expiring links to files.
Theka doesn't tie you to a provider: you use your own OpenAI, Anthropic, Gemini or Mistral key and decide where and how much it works.
An inexpensive model for invoices, a more capable one for contracts.
Basic, standard or full: more depth only where it is needed.
Spending limits for knowledge and for images and tables. Once the limit is reached, work resumes overnight.
Before bulk processing, see how many documents, how much time and how many tokens.
Security is in the way Theka is built, not in a separate module.
Each company has its own data space, separated from the others.
Files can live in your S3, SharePoint or MinIO. Export and delete whenever you like.
Search, chat and wiki only show what the user is allowed to open.
Who uploaded, approved, downloaded, signed. Always verifiable.
Datasheets, certificates, manuals and test reports for hundreds of product codes.
Every product has its own wiki: datasheet, certificate and manual are the cited sources.
Contracts, mandates and case files that pile up over the years and get lost every time someone changes role.
“When does the contract with customer X expire?” The answer comes with the clause and the page.
The same supplier seen by purchasing, quality and logistics.
One body of knowledge, different permissions: everyone sees what they need to.
A short definition to find your way among DAM, digital archive, semantic search, RAG and electronic signature.
Every answer and every sentence in the wiki carries a citation to the document and the page. If there is no source, the sentence is not written.
Search, chat and wiki always respect permissions: a user cannot find or read documents they would not be allowed to open.
Yes. Knowledge is a module: without it, Theka remains a complete document archive with versions, approvals, master data and signature.
In the EU. Files can stay in your own storage (S3, SharePoint, MinIO, SFTP) and you can export or delete them at any time.
You use your own AI provider key. You choose the model by category, set daily caps and see the estimate before any bulk processing.
For each document type you define the approval steps and who approves: everyone, at least one person or in order. Every decision (approved, rejected, changes requested) stays in the log.
Yes. Scanned pages go through OCR, images are described and data from tables and charts is extracted: everything becomes searchable and citable.
Yes, through the APIs: your PIM or ERP can upload and read documents, with permissions dedicated to the access key.
A guided demo with a sample of your archive.