Description:
Code-mixing, i.e., the mixing of two or more languages in a single utterance or conversation, is an extremely common phenomenon in multilingual societies. It is amply present in user-generated text, especially in social media. Therefore, CSS research that handles such text requires to process code-mixing; there are also interesting CSS and socio-linguistic questions around the phenomenon of code-mixing itself. In this tutorial, we will equip you with some basic tools and techniques for processing code-mixed text, starting with hands-on experiments with word-level language identification, all the way up to methods for building code-mixed text classifiers using massively multilingual language models.
Tutorial hosts: Monojit Choudhury & Sanad Rizvi
Slides: https://nlp-css-201-tutorials.github....
Code: https://colab.research.google.com/dri...
This is part of a larger tutorial series, NLP+CSS 201: Beyond the basics, which is organized by Ian Stewart and Katherine Keith. Website: https://nlp-css-201-tutorials.github....
On this page of the site you can watch the video online Tutorial 11: Processing Code-mixed Text with a duration of hours minute second in good quality, which was uploaded by the user NLP and CSS 201: Beyond the Basics 01 May 2022, share the link with friends and acquaintances, this video has already been watched 1,700 times on youtube and it was liked by 27 viewers. Enjoy your viewing!