Synthetic Bangla NLP research direction
I am developing a three-paper research cluster around a shared synthetic text generation pipeline for Bangla NLP.
The research will focus on:
- Dataset construction using LLM-based synthetic generation for low-resource settings.
- Evaluation of small and mid-size language models using shared Bangla benchmark settings.
- Applications of synthetic text data to computational social science research in Global South contexts.