Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

ChildDiffusion: Unlocking the Potential of Generative AI and Controllable Augmentations for Child Facial Data using Stable Diffusion and Large Language Models

Type de record:

paperdatasetmodel
Créateur:
FarYaoCor
Hôte:avatar
In this research work we have proposed high-level ChildDiffusion framework capable of generating photorealistic child facial samples and further embedding several intelligent augmentations on child facial data using short text prompts, detailed textual guidance from LLMs, and further image to image transformation using text guidance control conditioning thus providing an opportunity to curate fully synthetic large scale child datasets. The framework is validated by rendering high-quality child faces representing ethnicity data, micro expressions, face pose variations, eye blinking effects, facial accessories, different hair colours and styles, aging, multiple and different child gender subjects in a single frame. Addressing privacy concerns regarding child data acquisition requires a comprehensive approach that involves legal, ethical, and technological considerations. Keeping this in view this framework can be adapted to synthesise child facial data which can be effectively used for numerous downstream machine learning tasks. The proposed method circumvents common issues encountered in generative AI tools, such as temporal inconsistency and limited control over the rendered outputs. As an exemplary use case we have open-sourced child ethnicity data consisting of 2.5k child facial samples of five different classes which includes African, Asian, White, South Asian/ Indian, and Hispanic races by deploying the model in production inference phase. The rendered data undergoes rigorous qualitative as well as quantitative tests to cross validate its efficacy and further fine-tuning Yolo architecture for detecting and classifying child ethnicity as an exemplary downstream machine learning task. This work has been submitted to the IEEE Transactions Journal for possible publication

Visit

arxiv.org

Tasks

computer visionimage classification

Tags

Computer Vision and Pattern Recognition

Similaires

Mitigating Racial Biasness in Facial Recognition Technology Using Generative ModelsCendol: Open Instruction-tuned Generative Large Language Models for Indonesian LanguagesA Multilingual Sentiment Lexicon for Low-Resource Language Translation using Large Languages Models and Explainable AIMitigating Data Scarcity for Large Language ModelsExploring Multilingual Concepts of Human Value in Large Language Models: Is Value Alignment Consistent, Transferable and Controllable across Languages?Schema Generation for Large Knowledge Graphs Using Large Language Models

Mitigating Racial Biasness in Facial Recognition Technology Using Generative Models

In recent years, several researchers have worked to lessen biasness in face recognition systems sinc

Cendol: Open Instruction-tuned Generative Large Language Models for Indonesian Languages

Large language models (LLMs) show remarkable human-like capability in various domains and languages.

A Multilingual Sentiment Lexicon for Low-Resource Language Translation using Large Languages Models and Explainable AI

South Africa and the Democratic Republic of Congo (DRC) present a complex linguistic landscape with

Mitigating Data Scarcity for Large Language Models

In recent years, pretrained neural language models (PNLMs) have taken the field of natural language

Exploring Multilingual Concepts of Human Value in Large Language Models: Is Value Alignment Consistent, Transferable and Controllable across Languages?

Prior research has revealed that certain abstract concepts are linearly represented as directions in

Schema Generation for Large Knowledge Graphs Using Large Language Models

Schemas play a vital role in ensuring data quality and supporting usability in the Semantic Web and