NLP

Showing 1-3 of 3 results

CMU AI Repository – Names Corpus for NLP

Creators: Carnegie Mellon University (CMU) AI Repository
Publication Date: 1997-04-02
Creators: Carnegie Mellon University (CMU) AI Repository

The CMU AI Repository Names Corpus provides lists of male and female given names for use in natural language processing research. It is a widely used reference dataset for gender identification and name recognition tasks.

Creators: Steven Loria

TextBlob is an open-source Python library for processing textual data, providing simple APIs for common natural language processing (NLP) tasks such as sentiment analysis, part-of-speech tagging, noun phrase extraction, translation, and classification.

Popular Movies of TMDb

Creators: Mondal, Sankha Subhra
Publication Date: 2020
Creators: Mondal, Sankha Subhra

This dataset of the 10,000 most popular movies across the world has been fetched through the read API.
TMDB’s free API provides for developers and their team to programmatically fetch and use TMDb’s data.
Their API is to use as long as you attribute TMDb as the source of the data and/or images. Also, they update their API from time to time. The data set is 3.2 MB large. It offers valuable insights into global cinematic trends and preferences.

Each movie entry in the dataset includes the following attributes:

  • title: The name of the movie.
  • overview: A brief summary of the movie’s plot.
  • original_language: The language in which the movie was originally produced.
  • vote_average: The average user rating of the movie on TMDb.

Sign In

Register

Reset Password

Please enter your username or email address, you will receive a link to create a new password via email.