Expert insights on synthetic data

The latest

De-identifying and synthesizing healthcare PDFs of patient lab reports for model training and Expert Determination

PDFs are the hardest healthcare data to de-identify. Here is how Tonic Textual identifies PHI in lab reports and synthesizes realistic replacements for Expert Determination.

Blog posts

Building a RAG system on Databricks with your unstructured data using Tonic Textual

Technical deep dive
Technical deep dive
Tonic Textual

Using synthesized data for HIPAA expert determination

Healthcare
Healthcare
Data privacy
Data synthesis
Tonic Textual

Creating unstructured data pipelines for retrieval augmented generation

Product updates
Product updates
Tonic Textual

The challenges of preparing unstructured data for Generative AI

Technical deep dive
Technical deep dive
Tonic Textual

Tonic.ai product updates: July 2024

Product updates
Product updates
Tonic Structural
Tonic Textual

Sensitive data in text embeddings is recoverable

Technical deep dive
Technical deep dive
Tonic Textual

How to create de-identified embeddings with Tonic Textual & Pinecone

Technical deep dive
Technical deep dive
Tonic Textual

Tonic.ai product updates: May 2024

Product updates
Product updates
Tonic Structural
Tonic Textual

Tonic Textual available as Snowflake Native App to enable secure AI development

Product updates
Product updates
Tonic Textual

Achieving test data management totality: The difference between total coverage and "close enough"

Test data management
Test data management
Data de-identification
Tonic.ai editorial
Tonic Structural

Tonic.ai product updates: April 2024

Product updates
Product updates
Tonic Structural
Tonic Textual
Tonic Validate

How to de-identify legal documents with Tonic Textual

Data privacy
Data privacy
Tonic Textual