NameSuffixGenderClassifier

A simple Naive Bayes classifier that predicts gender from the last two letters of a name. This project trains a Multinomial Naive Bayes model using character bigrams and provides a command-line interface for predictions.

Project Overview

This project demonstrates a lightweight approach to gender classification based on name suffixes. It uses scikit-learn’s CountVectorizer with character bigrams and a MultinomialNB classifier. The implementation is kept intentionally minimal for educational purposes.

Dataset

The model expects a CSV file named genders.csv in the project root with at least the following columns:

name — the person’s name
gender — the target label (e.g., male, female)

Example:

name,gender Alex,male Maria,female

Method

Feature: last two letters of each name
Vectorization: character bigrams (ngram_range=(2,2))
Classifier: Multinomial Naive Bayes

Installation

Clone the repository
Install dependencies:

pip install pandas numpy scikit-learn

Usage

Ensure genders.csv is in the same directory as the script, then run:

python Gender_clasifiction_Naive_Bayes_Classifier.py

You will be prompted to enter a name. Type exit to quit.

Notes

This is a simple baseline model and may not perform well on diverse or international names.
For better results, consider richer features (full name, language-specific suffixes) and a larger dataset.

License

No license specified. Add a LICENSE file if you plan to share or reuse this project publicly.

Author

Created by mAhsanZafar.

Name		Name	Last commit message	Last commit date
Latest commit History 4 Commits
.gitignore		.gitignore
Gender_clasifiction_Naive_Bayes_Classifier.py		Gender_clasifiction_Naive_Bayes_Classifier.py
LICENSE		LICENSE
Male Female Classification using Naive Bayes Classifier.docx		Male Female Classification using Naive Bayes Classifier.docx
README.md		README.md
genders.csv		genders.csv

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

NameSuffixGenderClassifier

Project Overview

Dataset

Method

Installation

Usage

Notes

License

Author

About

Uh oh!

Releases

Packages

Uh oh!

Contributors

Uh oh!

Languages

Folders and files

Latest commit

History

Repository files navigation

NameSuffixGenderClassifier

Project Overview

Dataset

Method

Installation

Usage

Notes

License

Author

About

Topics

Resources

License

Uh oh!

Stars

Watchers

Forks

Releases

Packages 0

Uh oh!

Contributors

Uh oh!

Languages

Packages