Main Page: Difference between revisions

From LTRC
No edit summary
No edit summary
Line 163: Line 163:
     <div style="background: #f7fafc; border: 1px solid #edf2f7; padding: 15px 10px; border-radius: 10px;">
     <div style="background: #f7fafc; border: 1px solid #edf2f7; padding: 15px 10px; border-radius: 10px;">
       <div style="width: 50px; height: 50px; background: #3182ce; color: #ffffff; font-weight: bold; font-size: 1.1em; border-radius: 50%; display: flex; align-items: center; justify-content: center; margin: 0 auto 10px auto;">S</div>
       <div style="width: 50px; height: 50px; background: #3182ce; color: #ffffff; font-weight: bold; font-size: 1.1em; border-radius: 50%; display: flex; align-items: center; justify-content: center; margin: 0 auto 10px auto;">S</div>
       <strong style="color: #2d3748; display: block; font-size: 0.95em;">Intern 1</strong>
       <strong style="color: #2d3748; display: block; font-size: 0.95em;">Sourav Nath</strong>
       <span style="color: #a0aec0; font-size: 0.8em;">Contributor</span>
       <span style="color: #a0aec0; font-size: 0.8em;">Contributor</span>
     </div>
     </div>

Revision as of 12:42, 14 August 2026

Enhancing Scientific Wiki Articles in Indic Languages

Bridging the knowledge gap for the majority of India's population — making science, biodiversity, and innovation accessible in Hindi and Telugu.

RM
     Dr. Radhika Mamidi · Principal Investigator · LTRC, IIIT Hyderabad
 The Mission

Democratizing Scientific Knowledge

🌐

The Knowledge Gap

India records 117 million daily Wikipedia views, yet the vast majority of scientific content remains locked in English — out of reach for most Indian-language readers. We are fixing this at scale.

🇮🇳

Language Empowerment

Inspired by the success of native-language knowledge platforms elsewhere in the world, we are building a lasting Indian-language scientific corpus for generations to come.

🔬

Scientific Coverage

From biological sciences and chemistry to biodiversity, species taxonomy, and landmark inventions — every domain covered in both Hindi and Telugu.

🤖

AI-Powered Scale

Automated bot pipelines and NLP translation tools allow us to process WikiData and WikiSpecies at a scale impossible with manual effort alone.

Automated Technological Pipeline

     01
       🗄️ Data Sourcing

Harvesting structured data from WikiData and WikiSpecies — over 800,000 species entries.

     02
       🤖 Bot Creation

Automated Wikipedia bots programmed to generate, format, and upload article skeletons.

     03
       🔤 NLP Translation

Translation models fine-tuned on low-resource Indic language corpora.

     04
       ✅ Quality Review

Human-in-the-loop validation by LTRC linguists and domain experts before publication.

     05
       🌍 Open Release

All tools, datasets, and corpora released open-source for the global research community.

   Our Team

Contributing Interns

Recognizing the student researchers and contributors building and reviewing Indic scientific content.

K
     Krupal
     Contributor
A
     Abinaya
     Contributor
R
     Aditya  N Dwivedi
     Contributor
G
     Ameya Purohit
     Contributor
S
     Sourav Nath
     Contributor
R
     Intern 2 
     Contributor