All projects

Fair or Not? The Ethics of Unseen Data in AI Training

June 2024

An analysis of consent, transparency, ownership, and fairness in large-language-model training data.

AI EthicsLarge Language ModelsResearch Presentation

This presentation examined the ethical issues surrounding undisclosed or unseen data in artificial intelligence training, with a focus on large language models.

It considered questions of consent, transparency, ownership, and fairness in modern machine-learning datasets.

Highlights

  • Analyzed the ethical implications of using undisclosed data in LLM training.
  • Examined consent, transparency, ownership, and fairness in machine-learning datasets.