โ† Fine-Tuned Finance Models
Document AnalysisActive

TaxBERT

Community (Chalkidis et al.)

Legal-BERT fine-tuned on tax documents and EU directives โ€” the best open-source model for tax compliance document classification and European regulatory text analysis.

View on Hugging Face โ†’Read Paper โ†’
Base Model
RoBERTa-large
License
Apache-2.0
Downloads
300K+ monthly
Fine-Tuning
Fine-tuning on legal and tax documents
Training Data
12GB legal + tax documents, EU directives, tax regulations
Category
Document Analysis

TaxBERT is a domain-specific language model fine-tuned on legal and tax documents, built on the Legal-BERT architecture developed by Chalkidis and colleagues at the University of Athens. It is the most specialized open-source model available for tax-related natural language processing tasks.

The model was trained on a 12GB corpus of legal and tax documents including EU directives, tax regulations, VAT guidance, and corporate tax compliance materials. This extensive training on structured regulatory text gives TaxBERT a deep understanding of tax terminology, legal reasoning patterns, and the specific language used in tax legislation across multiple jurisdictions.

For tax compliance teams, accounting firms, and financial institutions dealing with cross-border tax obligations, TaxBERT provides a powerful tool for automating document classification, extracting key tax provisions, and analyzing regulatory text. Its Apache 2.0 license enables free commercial deployment, and its training on EU legal materials makes it particularly valuable for European financial institutions navigating the complex landscape of EU tax directives and member state regulations.

Finance Use Cases

  1. Tax document classification
  2. Tax regulation interpretation
  3. VAT and corporate tax document analysis
  4. Tax compliance document review
  5. EU tax directive analysis

Strengths

  • โœ“Best open-source model for tax and legal-financial document analysis
  • โœ“Trained on EU legal corpus โ€” ideal for European financial institutions
  • โœ“Apache 2.0 โ€” free commercial use

Limitations

  • โš Legal language focus โ€” not general finance
  • โš Older architecture vs modern LLMs
  • โš Less effective on unstructured financial documents

Fine-Tuning Details

Technique
Fine-tuning on legal and tax documents
Base Model
RoBERTa-large
Training Data
12GB legal + tax documents, EU directives, tax regulations
Developer
Community (Chalkidis et al.)

Explore in Finatune

Fine-Tuning Guides โ†’Fine-Tuning Use Cases โ†’Document Analysis Skills โ†’AI Models Directory โ†’

Related Models

Want to Fine-Tune Your Own Financial AI Model?

Start with our step-by-step guides covering LoRA, QLoRA, dataset preparation, and compliance frameworks for regulated financial institutions.

Browse Fine-Tuning Guides โ†’
โ† Back to Fine-Tuned Models Directory