Privacy Policy
How AI4Arunachal collects, manages, and protects contributor data and voice recordings.
Last updated: August 27, 2026
1 Information We Collect
To build authentic indigenous speech and translation datasets, we collect:
- Account Information: Name, username, and email address provided during registration.
- Linguistic Contributions: Audio recordings, dialect names, and text translations submitted via the platform.
- Usage & Analytics Data: Aggregated page interactions, reading time, and error diagnostics to improve system performance.
2 How We Use Your Data
Your contributed audio and text are used strictly for:
- Training automatic speech recognition (ASR), translation (NMT), and speech synthesis (TTS) models for indigenous and low-resource languages.
- Creating academic corpora accessible to indigenous communities, linguists, and researchers.
- Displaying aggregated contribution statistics and leaderboards on contributor dashboards.
3 Anonymization & Data Protection
Exported public datasets do not include personal emails or passwords. Audio samples are identified with randomized identifiers. We implement industry-standard encryption, CSRF protection, and secure server protocols to safeguard your credentials.
4 Your Rights & Data Removal
You retain the right to request deletion of your account and submitted audio recordings at any time by contacting our data management team at contact@ai4arunachal.in.