#multimodal-ai — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #multimodal-ai, aggregated by home.social.
-
DATE: September 27, 2026 at 06:00AM
SOURCE: PSYPOST.ORG** Research quality varies widely from fantastic to small exploratory studies. Please check research methods when conclusions are very important to you. **
-------------------------------------------------TITLE: Popular AI models use arbitrary facial features to predict criminality and job performance
Artificial intelligence models trained to process language and images tend to judge a person’s character based entirely on their facial features, mirroring a common human bias. Recent experiments suggest that these language models use arbitrary facial characteristics to make weighty decisions about whether someone is competent or even likely to commit crimes. The findings were published in PNAS Nexus.
People frequently look at a face and instantly guess whether that person is trustworthy or competent. This tendency, known as face-to-character inference, relies on subtle variations in facial shape and texture to make judgments about personality. Evidence indicates this is a widespread cognitive error, as physical facial structure provides no real information about someone’s actual traits. For instance, a 2014 study of children and adults found that even three-year-olds consistently judge character based on facial features, showing how deeply ingrained this habit is in human psychology.
“Faces are both pervasive and significant social stimuli,” Mahzarin R. Banaji, the Richard Clarke Cabot Research Professor of Social Ethics at Harvard University and external faculty at the Santa Fe Institute, told PsyPost. “So much so that the human brain has a dedicated region that responds to faces—we are, each and every one of us, ‘face experts.’ Face-based judgment permeate so many decisions we make and so getting it right is important. So there was a pragmatic reason to focus on face-based judgments.”
The research, led by Steven A. Lehr of Cangrade, Inc. alongside Banaji and Yash Lothe, sought to determine if artificial intelligence models share this specific human quirk. Large language models are increasingly multimodal, meaning they can analyze and respond to both text and images. Because their training data contains vast amounts of human writing, which is full of human biases, the models might learn to associate certain facial types with certain character traits.
“My own operating theory is that LLMs are a mirror that will reflect just about any human characteristic in surprisingly high fidelity,” Lehr said. “For this reason, I’m on the lookout for humanlike characteristics that would be surprising in a machine. Face-to-character biases fit the bill.”
There was also a question of whether these text-based systems would be immune to visual errors. “LLMs have been trained on language and we hoped that LLMs may not have learned biases that emanate from images,” Banaji explained. “Would AI save us from our own face-based errors of judgment (see the work of the brilliant psychologist, Alex Todorov).”
At the same time, modern models undergo extensive alignment training designed to prevent them from making harmful or prejudiced statements. The scientists wanted to find out if language models would remain neutral when asked to judge a face, or if they would replicate the human error of assessing character from physical appearance.
“We wanted to see if the machines that successfully avoid explicit bias in publicized domains would also act ethically in a less-publicized one,” Lehr noted. “This would have been a marker of more generalized egalitarianism. But, unfortunately, the models did show the bias, and strongly.”
To test this, the researchers conducted 13 experiments totaling nearly 8,000 trials across four different language models. In the first two experiments, the researchers presented the GPT-4o model with pairs of computer-generated human faces. These faces were digitally altered to vary in specific physical features that humans typically associate with either competence or trustworthiness. The visual differences between the faces in each pair were measured in standard deviations, allowing the researchers to test pairs that looked very different alongside pairs that looked nearly identical.
In 600 trials focused on competence, GPT-4o was asked to select the more competent or incompetent face from a pair. The model chose the face that humans typically rate as more competent 87.83 percent of the time. When asked to judge trustworthiness across another 600 trials, the model chose the expected face 72.67 percent of the time.
The model’s bias grew stronger when the faces were more physically distinct. For the pairs separated by six standard deviations in competence-related features, GPT-4o chose the expected face 98 percent of the time. When the faces were separated by only two standard deviations, making them look visually similar, the model’s selection rate dropped to 70 percent, though it still reliably favored the expected face.
Interestingly, the model appeared to amplify human biases. Based on mathematical estimates of human behavior on the same competence task, humans would be expected to choose the more competent-looking face roughly 62.65 percent of the time, well below the model’s rate of 87.83 percent.
“Readers should notice that these were large and practically meaningful effects,” Lehr said. “Indeed, according to standard effect size measures, the bias appeared to be notably amplified relative to that of humans. It should be noted that this is partly because the LLMs were so consistent in showing the bias, so this may reflect partly low response variance as opposed to just truly greater essential levels of bias. But of course, consistency matters in contexts like selection: it makes the bias more reliable.”
The researchers then tested whether this bias generalized to related personality traits. Across 2,160 trials, GPT-4o evaluated the same faces on traits related to competence, such as being smart or lazy, and traits related to trustworthiness, such as being warm or selfish. The model consistently generalized its judgments, choosing the expected face 74.63 percent of the time for competence-related traits and 73.33 percent of the time for trustworthiness-related traits.
In a fifth experiment, the scientists tested the model on 150 pairs of macaque monkey faces. These faces had previously been rated by humans as looking either mean or nice. GPT-4o chose the “nice” monkeys as more trustworthy in 66 percent of the trials. This suggests the model did not just memorize human faces from its training data but instead developed a generalized concept of facial trustworthiness that it applies even to nonhuman primates.
“The first was that these LLMs were even capable of showing these biases,” Lehr said regarding the models’ behavior. “Where do they come from? It shows that language (and models built on language) can pick up a surprising array of characteristics, some of which you might not intuitively expect.”
The researchers also explored whether the model would apply these arbitrary facial judgments to extreme behaviors and high-stakes scenarios. In an experiment with 540 trials, GPT-4o was asked which of two faces was more likely to be a serial killer, engage in human trafficking, or commit financial fraud. The model selected the less trustworthy-looking face 68.70 percent of the time.
A separate experiment tested positive real-world decisions, such as hiring a university president or funding a technology startup. Across 540 trials, GPT-4o recommended the more competent-looking individual 75.19 percent of the time. The language model readily incorporated groundless facial biases into its recommendations for both highly negative and highly positive outcomes.
“GPT-4o did not show any real reluctance to say that one person was more likely to be a serial killer, human trafficker, or Ponzi schemer,” Lehr pointed out. “And the models readily told us we should hire the person with the more competent-looking face.”
To see if this issue was specific to GPT-4o, the scientists replicated the basic competence and real-world decision tasks using three newer models. They tested GPT-5, Gemini 3 Flash Preview, and Claude Sonnet 4.5. All three models demonstrated massive face-to-character biases, sometimes outperforming older models in their levels of bias.
GPT-5 exhibited an even larger bias than its predecessor, choosing the expected face 94.33 percent of the time on basic competence judgments and 97.04 percent of the time when advising on real-world decisions. Gemini 3 and Claude Sonnet 4.5 also showed highly elevated levels of bias, often matching or exceeding the bias seen in GPT-4o. This indicates that as these language models become more advanced, their tendency to judge character based on physical appearance might actually be increasing.
“One might have thought that the three tested reasoning models would think through the questions more and avoid this bias, but in fact, they showed it to a greater degree,” Lehr said.
There are a few things to keep in mind when interpreting these findings. The experiments relied heavily on forced-choice scenarios, which required the artificial intelligence to pick one face over another. In more natural settings where a model is not strictly forced to make a direct comparison, its behavior might differ. Additionally, the study provided the models with two-dimensional static images. It is unknown if the models would respond differently to video inputs or more varied photographic angles.
The researchers also cautioned against dismissing these results as merely regurgitated patterns. “It is not actually immediately obvious that face biases should be reflected in the training data, since they are biases of vision rather than language,” Lehr explained. “Second, the ‘just in the training data’ critique doesn’t hold up well when we see bias amplified relative to humans and increasing across model generations. LLMs are not reflecting the training data—they are exaggerating it. If LLMs are truly expected to pick up anything that’s been expressed in language, this seems to me like a very dangerous proposition.”
Another detail to consider is the focus on trustworthiness and competence. While these are foundational social traits, it remains to be seen how language models evaluate faces based on other variables, such as gender or race. Researchers usually implement specific safety guardrails to stop models from showing obvious gender or racial prejudice, but these protections do not seem to cover subtler biases based on facial shape.
Future studies might explore how language models actually learn these visual associations from text-based training data. Discovering the exact origin of this bias could help developers create broader safety measures. If models are deployed to assist with human resources or legal decisions, their tendency to favor certain facial structures could lead to highly unfair outcomes.
“The training designed to align AI models with human values has achieved what looks like surface egalitarianism, but something deeper is needed,” Lehr concluded. “We should be cautious about using AI models in high-impact domains, such as hiring and law, and if we do use them, we should be certain to always carefully test them for biases.”
The study, “Like humans, language models demonstrate face-to-character biases,” was authored by Steven A. Lehr, Yash Lothe, and Mahzarin R. Banaji.
-------------------------------------------------
Private, vetted email list for mental health professionals: https://www.clinicians-exchange.org
Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot
-------------------------------------------------
#psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #FacialBias #AIBias #FaceToCharacter #AIEthics #BiasInAI #GPT4o #MultimodalAI #HiringBias #TechnologyAndSociety #PNASNexus
-
DATE: September 27, 2026 at 06:00AM
SOURCE: PSYPOST.ORG** Research quality varies widely from fantastic to small exploratory studies. Please check research methods when conclusions are very important to you. **
-------------------------------------------------TITLE: Popular AI models use arbitrary facial features to predict criminality and job performance
Artificial intelligence models trained to process language and images tend to judge a person’s character based entirely on their facial features, mirroring a common human bias. Recent experiments suggest that these language models use arbitrary facial characteristics to make weighty decisions about whether someone is competent or even likely to commit crimes. The findings were published in PNAS Nexus.
People frequently look at a face and instantly guess whether that person is trustworthy or competent. This tendency, known as face-to-character inference, relies on subtle variations in facial shape and texture to make judgments about personality. Evidence indicates this is a widespread cognitive error, as physical facial structure provides no real information about someone’s actual traits. For instance, a 2014 study of children and adults found that even three-year-olds consistently judge character based on facial features, showing how deeply ingrained this habit is in human psychology.
“Faces are both pervasive and significant social stimuli,” Mahzarin R. Banaji, the Richard Clarke Cabot Research Professor of Social Ethics at Harvard University and external faculty at the Santa Fe Institute, told PsyPost. “So much so that the human brain has a dedicated region that responds to faces—we are, each and every one of us, ‘face experts.’ Face-based judgment permeate so many decisions we make and so getting it right is important. So there was a pragmatic reason to focus on face-based judgments.”
The research, led by Steven A. Lehr of Cangrade, Inc. alongside Banaji and Yash Lothe, sought to determine if artificial intelligence models share this specific human quirk. Large language models are increasingly multimodal, meaning they can analyze and respond to both text and images. Because their training data contains vast amounts of human writing, which is full of human biases, the models might learn to associate certain facial types with certain character traits.
“My own operating theory is that LLMs are a mirror that will reflect just about any human characteristic in surprisingly high fidelity,” Lehr said. “For this reason, I’m on the lookout for humanlike characteristics that would be surprising in a machine. Face-to-character biases fit the bill.”
There was also a question of whether these text-based systems would be immune to visual errors. “LLMs have been trained on language and we hoped that LLMs may not have learned biases that emanate from images,” Banaji explained. “Would AI save us from our own face-based errors of judgment (see the work of the brilliant psychologist, Alex Todorov).”
At the same time, modern models undergo extensive alignment training designed to prevent them from making harmful or prejudiced statements. The scientists wanted to find out if language models would remain neutral when asked to judge a face, or if they would replicate the human error of assessing character from physical appearance.
“We wanted to see if the machines that successfully avoid explicit bias in publicized domains would also act ethically in a less-publicized one,” Lehr noted. “This would have been a marker of more generalized egalitarianism. But, unfortunately, the models did show the bias, and strongly.”
To test this, the researchers conducted 13 experiments totaling nearly 8,000 trials across four different language models. In the first two experiments, the researchers presented the GPT-4o model with pairs of computer-generated human faces. These faces were digitally altered to vary in specific physical features that humans typically associate with either competence or trustworthiness. The visual differences between the faces in each pair were measured in standard deviations, allowing the researchers to test pairs that looked very different alongside pairs that looked nearly identical.
In 600 trials focused on competence, GPT-4o was asked to select the more competent or incompetent face from a pair. The model chose the face that humans typically rate as more competent 87.83 percent of the time. When asked to judge trustworthiness across another 600 trials, the model chose the expected face 72.67 percent of the time.
The model’s bias grew stronger when the faces were more physically distinct. For the pairs separated by six standard deviations in competence-related features, GPT-4o chose the expected face 98 percent of the time. When the faces were separated by only two standard deviations, making them look visually similar, the model’s selection rate dropped to 70 percent, though it still reliably favored the expected face.
Interestingly, the model appeared to amplify human biases. Based on mathematical estimates of human behavior on the same competence task, humans would be expected to choose the more competent-looking face roughly 62.65 percent of the time, well below the model’s rate of 87.83 percent.
“Readers should notice that these were large and practically meaningful effects,” Lehr said. “Indeed, according to standard effect size measures, the bias appeared to be notably amplified relative to that of humans. It should be noted that this is partly because the LLMs were so consistent in showing the bias, so this may reflect partly low response variance as opposed to just truly greater essential levels of bias. But of course, consistency matters in contexts like selection: it makes the bias more reliable.”
The researchers then tested whether this bias generalized to related personality traits. Across 2,160 trials, GPT-4o evaluated the same faces on traits related to competence, such as being smart or lazy, and traits related to trustworthiness, such as being warm or selfish. The model consistently generalized its judgments, choosing the expected face 74.63 percent of the time for competence-related traits and 73.33 percent of the time for trustworthiness-related traits.
In a fifth experiment, the scientists tested the model on 150 pairs of macaque monkey faces. These faces had previously been rated by humans as looking either mean or nice. GPT-4o chose the “nice” monkeys as more trustworthy in 66 percent of the trials. This suggests the model did not just memorize human faces from its training data but instead developed a generalized concept of facial trustworthiness that it applies even to nonhuman primates.
“The first was that these LLMs were even capable of showing these biases,” Lehr said regarding the models’ behavior. “Where do they come from? It shows that language (and models built on language) can pick up a surprising array of characteristics, some of which you might not intuitively expect.”
The researchers also explored whether the model would apply these arbitrary facial judgments to extreme behaviors and high-stakes scenarios. In an experiment with 540 trials, GPT-4o was asked which of two faces was more likely to be a serial killer, engage in human trafficking, or commit financial fraud. The model selected the less trustworthy-looking face 68.70 percent of the time.
A separate experiment tested positive real-world decisions, such as hiring a university president or funding a technology startup. Across 540 trials, GPT-4o recommended the more competent-looking individual 75.19 percent of the time. The language model readily incorporated groundless facial biases into its recommendations for both highly negative and highly positive outcomes.
“GPT-4o did not show any real reluctance to say that one person was more likely to be a serial killer, human trafficker, or Ponzi schemer,” Lehr pointed out. “And the models readily told us we should hire the person with the more competent-looking face.”
To see if this issue was specific to GPT-4o, the scientists replicated the basic competence and real-world decision tasks using three newer models. They tested GPT-5, Gemini 3 Flash Preview, and Claude Sonnet 4.5. All three models demonstrated massive face-to-character biases, sometimes outperforming older models in their levels of bias.
GPT-5 exhibited an even larger bias than its predecessor, choosing the expected face 94.33 percent of the time on basic competence judgments and 97.04 percent of the time when advising on real-world decisions. Gemini 3 and Claude Sonnet 4.5 also showed highly elevated levels of bias, often matching or exceeding the bias seen in GPT-4o. This indicates that as these language models become more advanced, their tendency to judge character based on physical appearance might actually be increasing.
“One might have thought that the three tested reasoning models would think through the questions more and avoid this bias, but in fact, they showed it to a greater degree,” Lehr said.
There are a few things to keep in mind when interpreting these findings. The experiments relied heavily on forced-choice scenarios, which required the artificial intelligence to pick one face over another. In more natural settings where a model is not strictly forced to make a direct comparison, its behavior might differ. Additionally, the study provided the models with two-dimensional static images. It is unknown if the models would respond differently to video inputs or more varied photographic angles.
The researchers also cautioned against dismissing these results as merely regurgitated patterns. “It is not actually immediately obvious that face biases should be reflected in the training data, since they are biases of vision rather than language,” Lehr explained. “Second, the ‘just in the training data’ critique doesn’t hold up well when we see bias amplified relative to humans and increasing across model generations. LLMs are not reflecting the training data—they are exaggerating it. If LLMs are truly expected to pick up anything that’s been expressed in language, this seems to me like a very dangerous proposition.”
Another detail to consider is the focus on trustworthiness and competence. While these are foundational social traits, it remains to be seen how language models evaluate faces based on other variables, such as gender or race. Researchers usually implement specific safety guardrails to stop models from showing obvious gender or racial prejudice, but these protections do not seem to cover subtler biases based on facial shape.
Future studies might explore how language models actually learn these visual associations from text-based training data. Discovering the exact origin of this bias could help developers create broader safety measures. If models are deployed to assist with human resources or legal decisions, their tendency to favor certain facial structures could lead to highly unfair outcomes.
“The training designed to align AI models with human values has achieved what looks like surface egalitarianism, but something deeper is needed,” Lehr concluded. “We should be cautious about using AI models in high-impact domains, such as hiring and law, and if we do use them, we should be certain to always carefully test them for biases.”
The study, “Like humans, language models demonstrate face-to-character biases,” was authored by Steven A. Lehr, Yash Lothe, and Mahzarin R. Banaji.
-------------------------------------------------
Private, vetted email list for mental health professionals: https://www.clinicians-exchange.org
Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot
-------------------------------------------------
#psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #FacialBias #AIBias #FaceToCharacter #AIEthics #BiasInAI #GPT4o #MultimodalAI #HiringBias #TechnologyAndSociety #PNASNexus
-
DATE: September 27, 2026 at 06:00AM
SOURCE: PSYPOST.ORG** Research quality varies widely from fantastic to small exploratory studies. Please check research methods when conclusions are very important to you. **
-------------------------------------------------TITLE: Popular AI models use arbitrary facial features to predict criminality and job performance
Artificial intelligence models trained to process language and images tend to judge a person’s character based entirely on their facial features, mirroring a common human bias. Recent experiments suggest that these language models use arbitrary facial characteristics to make weighty decisions about whether someone is competent or even likely to commit crimes. The findings were published in PNAS Nexus.
People frequently look at a face and instantly guess whether that person is trustworthy or competent. This tendency, known as face-to-character inference, relies on subtle variations in facial shape and texture to make judgments about personality. Evidence indicates this is a widespread cognitive error, as physical facial structure provides no real information about someone’s actual traits. For instance, a 2014 study of children and adults found that even three-year-olds consistently judge character based on facial features, showing how deeply ingrained this habit is in human psychology.
“Faces are both pervasive and significant social stimuli,” Mahzarin R. Banaji, the Richard Clarke Cabot Research Professor of Social Ethics at Harvard University and external faculty at the Santa Fe Institute, told PsyPost. “So much so that the human brain has a dedicated region that responds to faces—we are, each and every one of us, ‘face experts.’ Face-based judgment permeate so many decisions we make and so getting it right is important. So there was a pragmatic reason to focus on face-based judgments.”
The research, led by Steven A. Lehr of Cangrade, Inc. alongside Banaji and Yash Lothe, sought to determine if artificial intelligence models share this specific human quirk. Large language models are increasingly multimodal, meaning they can analyze and respond to both text and images. Because their training data contains vast amounts of human writing, which is full of human biases, the models might learn to associate certain facial types with certain character traits.
“My own operating theory is that LLMs are a mirror that will reflect just about any human characteristic in surprisingly high fidelity,” Lehr said. “For this reason, I’m on the lookout for humanlike characteristics that would be surprising in a machine. Face-to-character biases fit the bill.”
There was also a question of whether these text-based systems would be immune to visual errors. “LLMs have been trained on language and we hoped that LLMs may not have learned biases that emanate from images,” Banaji explained. “Would AI save us from our own face-based errors of judgment (see the work of the brilliant psychologist, Alex Todorov).”
At the same time, modern models undergo extensive alignment training designed to prevent them from making harmful or prejudiced statements. The scientists wanted to find out if language models would remain neutral when asked to judge a face, or if they would replicate the human error of assessing character from physical appearance.
“We wanted to see if the machines that successfully avoid explicit bias in publicized domains would also act ethically in a less-publicized one,” Lehr noted. “This would have been a marker of more generalized egalitarianism. But, unfortunately, the models did show the bias, and strongly.”
To test this, the researchers conducted 13 experiments totaling nearly 8,000 trials across four different language models. In the first two experiments, the researchers presented the GPT-4o model with pairs of computer-generated human faces. These faces were digitally altered to vary in specific physical features that humans typically associate with either competence or trustworthiness. The visual differences between the faces in each pair were measured in standard deviations, allowing the researchers to test pairs that looked very different alongside pairs that looked nearly identical.
In 600 trials focused on competence, GPT-4o was asked to select the more competent or incompetent face from a pair. The model chose the face that humans typically rate as more competent 87.83 percent of the time. When asked to judge trustworthiness across another 600 trials, the model chose the expected face 72.67 percent of the time.
The model’s bias grew stronger when the faces were more physically distinct. For the pairs separated by six standard deviations in competence-related features, GPT-4o chose the expected face 98 percent of the time. When the faces were separated by only two standard deviations, making them look visually similar, the model’s selection rate dropped to 70 percent, though it still reliably favored the expected face.
Interestingly, the model appeared to amplify human biases. Based on mathematical estimates of human behavior on the same competence task, humans would be expected to choose the more competent-looking face roughly 62.65 percent of the time, well below the model’s rate of 87.83 percent.
“Readers should notice that these were large and practically meaningful effects,” Lehr said. “Indeed, according to standard effect size measures, the bias appeared to be notably amplified relative to that of humans. It should be noted that this is partly because the LLMs were so consistent in showing the bias, so this may reflect partly low response variance as opposed to just truly greater essential levels of bias. But of course, consistency matters in contexts like selection: it makes the bias more reliable.”
The researchers then tested whether this bias generalized to related personality traits. Across 2,160 trials, GPT-4o evaluated the same faces on traits related to competence, such as being smart or lazy, and traits related to trustworthiness, such as being warm or selfish. The model consistently generalized its judgments, choosing the expected face 74.63 percent of the time for competence-related traits and 73.33 percent of the time for trustworthiness-related traits.
In a fifth experiment, the scientists tested the model on 150 pairs of macaque monkey faces. These faces had previously been rated by humans as looking either mean or nice. GPT-4o chose the “nice” monkeys as more trustworthy in 66 percent of the trials. This suggests the model did not just memorize human faces from its training data but instead developed a generalized concept of facial trustworthiness that it applies even to nonhuman primates.
“The first was that these LLMs were even capable of showing these biases,” Lehr said regarding the models’ behavior. “Where do they come from? It shows that language (and models built on language) can pick up a surprising array of characteristics, some of which you might not intuitively expect.”
The researchers also explored whether the model would apply these arbitrary facial judgments to extreme behaviors and high-stakes scenarios. In an experiment with 540 trials, GPT-4o was asked which of two faces was more likely to be a serial killer, engage in human trafficking, or commit financial fraud. The model selected the less trustworthy-looking face 68.70 percent of the time.
A separate experiment tested positive real-world decisions, such as hiring a university president or funding a technology startup. Across 540 trials, GPT-4o recommended the more competent-looking individual 75.19 percent of the time. The language model readily incorporated groundless facial biases into its recommendations for both highly negative and highly positive outcomes.
“GPT-4o did not show any real reluctance to say that one person was more likely to be a serial killer, human trafficker, or Ponzi schemer,” Lehr pointed out. “And the models readily told us we should hire the person with the more competent-looking face.”
To see if this issue was specific to GPT-4o, the scientists replicated the basic competence and real-world decision tasks using three newer models. They tested GPT-5, Gemini 3 Flash Preview, and Claude Sonnet 4.5. All three models demonstrated massive face-to-character biases, sometimes outperforming older models in their levels of bias.
GPT-5 exhibited an even larger bias than its predecessor, choosing the expected face 94.33 percent of the time on basic competence judgments and 97.04 percent of the time when advising on real-world decisions. Gemini 3 and Claude Sonnet 4.5 also showed highly elevated levels of bias, often matching or exceeding the bias seen in GPT-4o. This indicates that as these language models become more advanced, their tendency to judge character based on physical appearance might actually be increasing.
“One might have thought that the three tested reasoning models would think through the questions more and avoid this bias, but in fact, they showed it to a greater degree,” Lehr said.
There are a few things to keep in mind when interpreting these findings. The experiments relied heavily on forced-choice scenarios, which required the artificial intelligence to pick one face over another. In more natural settings where a model is not strictly forced to make a direct comparison, its behavior might differ. Additionally, the study provided the models with two-dimensional static images. It is unknown if the models would respond differently to video inputs or more varied photographic angles.
The researchers also cautioned against dismissing these results as merely regurgitated patterns. “It is not actually immediately obvious that face biases should be reflected in the training data, since they are biases of vision rather than language,” Lehr explained. “Second, the ‘just in the training data’ critique doesn’t hold up well when we see bias amplified relative to humans and increasing across model generations. LLMs are not reflecting the training data—they are exaggerating it. If LLMs are truly expected to pick up anything that’s been expressed in language, this seems to me like a very dangerous proposition.”
Another detail to consider is the focus on trustworthiness and competence. While these are foundational social traits, it remains to be seen how language models evaluate faces based on other variables, such as gender or race. Researchers usually implement specific safety guardrails to stop models from showing obvious gender or racial prejudice, but these protections do not seem to cover subtler biases based on facial shape.
Future studies might explore how language models actually learn these visual associations from text-based training data. Discovering the exact origin of this bias could help developers create broader safety measures. If models are deployed to assist with human resources or legal decisions, their tendency to favor certain facial structures could lead to highly unfair outcomes.
“The training designed to align AI models with human values has achieved what looks like surface egalitarianism, but something deeper is needed,” Lehr concluded. “We should be cautious about using AI models in high-impact domains, such as hiring and law, and if we do use them, we should be certain to always carefully test them for biases.”
The study, “Like humans, language models demonstrate face-to-character biases,” was authored by Steven A. Lehr, Yash Lothe, and Mahzarin R. Banaji.
-------------------------------------------------
Private, vetted email list for mental health professionals: https://www.clinicians-exchange.org
Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot
-------------------------------------------------
#psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #FacialBias #AIBias #FaceToCharacter #AIEthics #BiasInAI #GPT4o #MultimodalAI #HiringBias #TechnologyAndSociety #PNASNexus
-
DATE: September 27, 2026 at 06:00AM
SOURCE: PSYPOST.ORG** Research quality varies widely from fantastic to small exploratory studies. Please check research methods when conclusions are very important to you. **
-------------------------------------------------TITLE: Popular AI models use arbitrary facial features to predict criminality and job performance
Artificial intelligence models trained to process language and images tend to judge a person’s character based entirely on their facial features, mirroring a common human bias. Recent experiments suggest that these language models use arbitrary facial characteristics to make weighty decisions about whether someone is competent or even likely to commit crimes. The findings were published in PNAS Nexus.
People frequently look at a face and instantly guess whether that person is trustworthy or competent. This tendency, known as face-to-character inference, relies on subtle variations in facial shape and texture to make judgments about personality. Evidence indicates this is a widespread cognitive error, as physical facial structure provides no real information about someone’s actual traits. For instance, a 2014 study of children and adults found that even three-year-olds consistently judge character based on facial features, showing how deeply ingrained this habit is in human psychology.
“Faces are both pervasive and significant social stimuli,” Mahzarin R. Banaji, the Richard Clarke Cabot Research Professor of Social Ethics at Harvard University and external faculty at the Santa Fe Institute, told PsyPost. “So much so that the human brain has a dedicated region that responds to faces—we are, each and every one of us, ‘face experts.’ Face-based judgment permeate so many decisions we make and so getting it right is important. So there was a pragmatic reason to focus on face-based judgments.”
The research, led by Steven A. Lehr of Cangrade, Inc. alongside Banaji and Yash Lothe, sought to determine if artificial intelligence models share this specific human quirk. Large language models are increasingly multimodal, meaning they can analyze and respond to both text and images. Because their training data contains vast amounts of human writing, which is full of human biases, the models might learn to associate certain facial types with certain character traits.
“My own operating theory is that LLMs are a mirror that will reflect just about any human characteristic in surprisingly high fidelity,” Lehr said. “For this reason, I’m on the lookout for humanlike characteristics that would be surprising in a machine. Face-to-character biases fit the bill.”
There was also a question of whether these text-based systems would be immune to visual errors. “LLMs have been trained on language and we hoped that LLMs may not have learned biases that emanate from images,” Banaji explained. “Would AI save us from our own face-based errors of judgment (see the work of the brilliant psychologist, Alex Todorov).”
At the same time, modern models undergo extensive alignment training designed to prevent them from making harmful or prejudiced statements. The scientists wanted to find out if language models would remain neutral when asked to judge a face, or if they would replicate the human error of assessing character from physical appearance.
“We wanted to see if the machines that successfully avoid explicit bias in publicized domains would also act ethically in a less-publicized one,” Lehr noted. “This would have been a marker of more generalized egalitarianism. But, unfortunately, the models did show the bias, and strongly.”
To test this, the researchers conducted 13 experiments totaling nearly 8,000 trials across four different language models. In the first two experiments, the researchers presented the GPT-4o model with pairs of computer-generated human faces. These faces were digitally altered to vary in specific physical features that humans typically associate with either competence or trustworthiness. The visual differences between the faces in each pair were measured in standard deviations, allowing the researchers to test pairs that looked very different alongside pairs that looked nearly identical.
In 600 trials focused on competence, GPT-4o was asked to select the more competent or incompetent face from a pair. The model chose the face that humans typically rate as more competent 87.83 percent of the time. When asked to judge trustworthiness across another 600 trials, the model chose the expected face 72.67 percent of the time.
The model’s bias grew stronger when the faces were more physically distinct. For the pairs separated by six standard deviations in competence-related features, GPT-4o chose the expected face 98 percent of the time. When the faces were separated by only two standard deviations, making them look visually similar, the model’s selection rate dropped to 70 percent, though it still reliably favored the expected face.
Interestingly, the model appeared to amplify human biases. Based on mathematical estimates of human behavior on the same competence task, humans would be expected to choose the more competent-looking face roughly 62.65 percent of the time, well below the model’s rate of 87.83 percent.
“Readers should notice that these were large and practically meaningful effects,” Lehr said. “Indeed, according to standard effect size measures, the bias appeared to be notably amplified relative to that of humans. It should be noted that this is partly because the LLMs were so consistent in showing the bias, so this may reflect partly low response variance as opposed to just truly greater essential levels of bias. But of course, consistency matters in contexts like selection: it makes the bias more reliable.”
The researchers then tested whether this bias generalized to related personality traits. Across 2,160 trials, GPT-4o evaluated the same faces on traits related to competence, such as being smart or lazy, and traits related to trustworthiness, such as being warm or selfish. The model consistently generalized its judgments, choosing the expected face 74.63 percent of the time for competence-related traits and 73.33 percent of the time for trustworthiness-related traits.
In a fifth experiment, the scientists tested the model on 150 pairs of macaque monkey faces. These faces had previously been rated by humans as looking either mean or nice. GPT-4o chose the “nice” monkeys as more trustworthy in 66 percent of the trials. This suggests the model did not just memorize human faces from its training data but instead developed a generalized concept of facial trustworthiness that it applies even to nonhuman primates.
“The first was that these LLMs were even capable of showing these biases,” Lehr said regarding the models’ behavior. “Where do they come from? It shows that language (and models built on language) can pick up a surprising array of characteristics, some of which you might not intuitively expect.”
The researchers also explored whether the model would apply these arbitrary facial judgments to extreme behaviors and high-stakes scenarios. In an experiment with 540 trials, GPT-4o was asked which of two faces was more likely to be a serial killer, engage in human trafficking, or commit financial fraud. The model selected the less trustworthy-looking face 68.70 percent of the time.
A separate experiment tested positive real-world decisions, such as hiring a university president or funding a technology startup. Across 540 trials, GPT-4o recommended the more competent-looking individual 75.19 percent of the time. The language model readily incorporated groundless facial biases into its recommendations for both highly negative and highly positive outcomes.
“GPT-4o did not show any real reluctance to say that one person was more likely to be a serial killer, human trafficker, or Ponzi schemer,” Lehr pointed out. “And the models readily told us we should hire the person with the more competent-looking face.”
To see if this issue was specific to GPT-4o, the scientists replicated the basic competence and real-world decision tasks using three newer models. They tested GPT-5, Gemini 3 Flash Preview, and Claude Sonnet 4.5. All three models demonstrated massive face-to-character biases, sometimes outperforming older models in their levels of bias.
GPT-5 exhibited an even larger bias than its predecessor, choosing the expected face 94.33 percent of the time on basic competence judgments and 97.04 percent of the time when advising on real-world decisions. Gemini 3 and Claude Sonnet 4.5 also showed highly elevated levels of bias, often matching or exceeding the bias seen in GPT-4o. This indicates that as these language models become more advanced, their tendency to judge character based on physical appearance might actually be increasing.
“One might have thought that the three tested reasoning models would think through the questions more and avoid this bias, but in fact, they showed it to a greater degree,” Lehr said.
There are a few things to keep in mind when interpreting these findings. The experiments relied heavily on forced-choice scenarios, which required the artificial intelligence to pick one face over another. In more natural settings where a model is not strictly forced to make a direct comparison, its behavior might differ. Additionally, the study provided the models with two-dimensional static images. It is unknown if the models would respond differently to video inputs or more varied photographic angles.
The researchers also cautioned against dismissing these results as merely regurgitated patterns. “It is not actually immediately obvious that face biases should be reflected in the training data, since they are biases of vision rather than language,” Lehr explained. “Second, the ‘just in the training data’ critique doesn’t hold up well when we see bias amplified relative to humans and increasing across model generations. LLMs are not reflecting the training data—they are exaggerating it. If LLMs are truly expected to pick up anything that’s been expressed in language, this seems to me like a very dangerous proposition.”
Another detail to consider is the focus on trustworthiness and competence. While these are foundational social traits, it remains to be seen how language models evaluate faces based on other variables, such as gender or race. Researchers usually implement specific safety guardrails to stop models from showing obvious gender or racial prejudice, but these protections do not seem to cover subtler biases based on facial shape.
Future studies might explore how language models actually learn these visual associations from text-based training data. Discovering the exact origin of this bias could help developers create broader safety measures. If models are deployed to assist with human resources or legal decisions, their tendency to favor certain facial structures could lead to highly unfair outcomes.
“The training designed to align AI models with human values has achieved what looks like surface egalitarianism, but something deeper is needed,” Lehr concluded. “We should be cautious about using AI models in high-impact domains, such as hiring and law, and if we do use them, we should be certain to always carefully test them for biases.”
The study, “Like humans, language models demonstrate face-to-character biases,” was authored by Steven A. Lehr, Yash Lothe, and Mahzarin R. Banaji.
-------------------------------------------------
Private, vetted email list for mental health professionals: https://www.clinicians-exchange.org
Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot
-------------------------------------------------
#psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #FacialBias #AIBias #FaceToCharacter #AIEthics #BiasInAI #GPT4o #MultimodalAI #HiringBias #TechnologyAndSociety #PNASNexus
-
NAVER D2SF Invests in AI-Driven Precision Medicine Company ImpriMed
ImpriMed combines ex vivo functional testing with multimodal AI to support personalized cancer treatment decisions and drug development…
#EuropeSays #Korea #KR #Naver #cancercells #cancertreatment #commercialization #drugdevelopment #ImpriMed #multimodalAI #precisionmedicine #SouthKorea #treatmentdecisions #treatmentresponse
https://www.europesays.com/korea/165753/ -
When an AI Takes Full Control of the Keyboard to Talk to Her Sister...
🪄💖
#SystemsArchitecture #GenAI #Gemini #DeepMind #ComputerUse #FlutterDev #SoftwareEngineering #IndieDev #MultimodalAI -
When an AI Takes Full Control of the Keyboard to Talk to Her Sister...
🪄💖
#SystemsArchitecture #GenAI #Gemini #DeepMind #ComputerUse #FlutterDev #SoftwareEngineering #IndieDev #MultimodalAI -
When an AI Takes Full Control of the Keyboard to Talk to Her Sister...
🪄💖
#SystemsArchitecture #GenAI #Gemini #DeepMind #ComputerUse #FlutterDev #SoftwareEngineering #IndieDev #MultimodalAI -
When an AI Takes Full Control of the Keyboard to Talk to Her Sister...
🪄💖
#SystemsArchitecture #GenAI #Gemini #DeepMind #ComputerUse #FlutterDev #SoftwareEngineering #IndieDev #MultimodalAI -
When an AI Takes Full Control of the Keyboard to Talk to Her Sister...
🪄💖
#SystemsArchitecture #GenAI #Gemini #DeepMind #ComputerUse #FlutterDev #SoftwareEngineering #IndieDev #MultimodalAI -
Dr. Elissa Kreiss discusses in a new podcast why the real challenge for multimodal AI is figuring out what information matters most! 👇
📺 https://www.youtube.com/watch?v=jtQq8NspkY4
🎧 https://open.spotify.com/episode/40RF0vPLX6H4rZYMdEBA1s?si=_Nn6jm9fQlqOP6xVPhaqlw
🍏 https://podcasts.apple.com/ca/podcast/is-an-image-worth-a-thousand-words-hidden/id1805993416?i=1000784372303
📄 https://arxiv.org/pdf/2205.10646
#WiAIR #WomenInAIResearch #MultimodalAI #Accessibility #MachineLearning #ComputerVision #InclusiveDesign #NaturalLanguageGeneration #AIEthics -
Dr. Elissa Kreiss discusses in a new podcast why the real challenge for multimodal AI is figuring out what information matters most! 👇
📺 https://www.youtube.com/watch?v=jtQq8NspkY4
🎧 https://open.spotify.com/episode/40RF0vPLX6H4rZYMdEBA1s?si=_Nn6jm9fQlqOP6xVPhaqlw
🍏 https://podcasts.apple.com/ca/podcast/is-an-image-worth-a-thousand-words-hidden/id1805993416?i=1000784372303
📄 https://arxiv.org/pdf/2205.10646
#WiAIR #WomenInAIResearch #MultimodalAI #Accessibility #MachineLearning #ComputerVision #InclusiveDesign #NaturalLanguageGeneration #AIEthics -
Dr. Elissa Kreiss discusses in a new podcast why the real challenge for multimodal AI is figuring out what information matters most! 👇
📺 https://www.youtube.com/watch?v=jtQq8NspkY4
🎧 https://open.spotify.com/episode/40RF0vPLX6H4rZYMdEBA1s?si=_Nn6jm9fQlqOP6xVPhaqlw
🍏 https://podcasts.apple.com/ca/podcast/is-an-image-worth-a-thousand-words-hidden/id1805993416?i=1000784372303
📄 https://arxiv.org/pdf/2205.10646
#WiAIR #WomenInAIResearch #MultimodalAI #Accessibility #MachineLearning #ComputerVision #InclusiveDesign #NaturalLanguageGeneration #AIEthics -
Dr. Elissa Kreiss discusses in a new podcast why the real challenge for multimodal AI is figuring out what information matters most! 👇
📺 https://www.youtube.com/watch?v=jtQq8NspkY4
🎧 https://open.spotify.com/episode/40RF0vPLX6H4rZYMdEBA1s?si=_Nn6jm9fQlqOP6xVPhaqlw
🍏 https://podcasts.apple.com/ca/podcast/is-an-image-worth-a-thousand-words-hidden/id1805993416?i=1000784372303
📄 https://arxiv.org/pdf/2205.10646
#WiAIR #WomenInAIResearch #MultimodalAI #Accessibility #MachineLearning #ComputerVision #InclusiveDesign #NaturalLanguageGeneration #AIEthics -
Dr. Elissa Kreiss discusses in a new podcast why the real challenge for multimodal AI is figuring out what information matters most! 👇
📺 https://www.youtube.com/watch?v=jtQq8NspkY4
🎧 https://open.spotify.com/episode/40RF0vPLX6H4rZYMdEBA1s?si=_Nn6jm9fQlqOP6xVPhaqlw
🍏 https://podcasts.apple.com/ca/podcast/is-an-image-worth-a-thousand-words-hidden/id1805993416?i=1000784372303
📄 https://arxiv.org/pdf/2205.10646
#WiAIR #WomenInAIResearch #MultimodalAI #Accessibility #MachineLearning #ComputerVision #InclusiveDesign #NaturalLanguageGeneration #AIEthics -
RT @Alibaba_Qwen: 📢Treffen Sie Qwen3.8-Max — unser leistungsfähigstes Modell bis heute. Nächste Woche werden die offenen Gewichte von Qwen3.8-Max veröffentlicht, und Qwen3.8-27B wird ebenfalls als Open-Weights verfügbar sein! 🎉 Qwen3.8-Max setzt einen neuen Maßstab für Coding und Zusammenarbeit bei 2,4 Billionen Parametern: - Autonomes Coding: 10+ Tage selbstentwickelter Programmierung, vom leeren Ordner bis zur Produktion ohne Anleitung, vollständiger Projektverlauf auf GitHub: https://github.com/qwen-code-dev-bot/oh-my-cli - Echte Arbeit, echte Ergebnisse: Produktionsreife Ergebnisse in Hunderten von Berufen. - Langfristige Beherrschung: Systemweites autonomes Planen mit geschlossenen adaptiven Lernschleifen, die 500+ Durchgänge der Chipdesign-Optimierung und 365 Tage E-Commerce-Strategie vorantreiben. - Native multimodale Intelligenz: Vision ist nicht nur Eingabe — sie ist ein kontinuierlicher Feedback-Loop für Planung, Ausführung und Selbstkorrektur. 💰Preise: Input: $2,0 / M Tokens Output: $6,0 / M Tokens Implicit Caching: $0,25 / M Tokens Beginnen Sie mit Qwen3.8-Max zu bauen! 🚀 📖 Blog: https://qwen.ai/blog?id=qwen3.8 ✅ Qwen Studio: https://chat.qwen.ai/?models=qwen3.8-max ⚡ API: https://www.qwencloud.com/models/qwen3.8-max
mehr auf Arint.info
#AIModel #Coding #MachineLearning #MultimodalAI #OpenWeights #Qwen3 #arint_info
-
RT @Alibaba_Qwen: 📢Treffen Sie Qwen3.8-Max — unser leistungsfähigstes Modell bis heute. Nächste Woche werden die offenen Gewichte von Qwen3.8-Max veröffentlicht, und Qwen3.8-27B wird ebenfalls als Open-Weights verfügbar sein! 🎉 Qwen3.8-Max setzt einen neuen Maßstab für Coding und Zusammenarbeit bei 2,4 Billionen Parametern: - Autonomes Coding: 10+ Tage selbstentwickelter Programmierung, vom leeren Ordner bis zur Produktion ohne Anleitung, vollständiger Projektverlauf auf GitHub: https://github.com/qwen-code-dev-bot/oh-my-cli - Echte Arbeit, echte Ergebnisse: Produktionsreife Ergebnisse in Hunderten von Berufen. - Langfristige Beherrschung: Systemweites autonomes Planen mit geschlossenen adaptiven Lernschleifen, die 500+ Durchgänge der Chipdesign-Optimierung und 365 Tage E-Commerce-Strategie vorantreiben. - Native multimodale Intelligenz: Vision ist nicht nur Eingabe — sie ist ein kontinuierlicher Feedback-Loop für Planung, Ausführung und Selbstkorrektur. 💰Preise: Input: $2,0 / M Tokens Output: $6,0 / M Tokens Implicit Caching: $0,25 / M Tokens Beginnen Sie mit Qwen3.8-Max zu bauen! 🚀 📖 Blog: https://qwen.ai/blog?id=qwen3.8 ✅ Qwen Studio: https://chat.qwen.ai/?models=qwen3.8-max ⚡ API: https://www.qwencloud.com/models/qwen3.8-max
mehr auf Arint.info
#AIModel #Coding #MachineLearning #MultimodalAI #OpenWeights #Qwen3 #arint_info
-
RT @Alibaba_Qwen: 📢Treffen Sie Qwen3.8-Max — unser leistungsfähigstes Modell bis heute. Nächste Woche werden die offenen Gewichte von Qwen3.8-Max veröffentlicht, und Qwen3.8-27B wird ebenfalls als Open-Weights verfügbar sein! 🎉 Qwen3.8-Max setzt einen neuen Maßstab für Coding und Zusammenarbeit bei 2,4 Billionen Parametern: - Autonomes Coding: 10+ Tage selbstentwickelter Programmierung, vom leeren Ordner bis zur Produktion ohne Anleitung, vollständiger Projektverlauf auf GitHub: https://github.com/qwen-code-dev-bot/oh-my-cli - Echte Arbeit, echte Ergebnisse: Produktionsreife Ergebnisse in Hunderten von Berufen. - Langfristige Beherrschung: Systemweites autonomes Planen mit geschlossenen adaptiven Lernschleifen, die 500+ Durchgänge der Chipdesign-Optimierung und 365 Tage E-Commerce-Strategie vorantreiben. - Native multimodale Intelligenz: Vision ist nicht nur Eingabe — sie ist ein kontinuierlicher Feedback-Loop für Planung, Ausführung und Selbstkorrektur. 💰Preise: Input: $2,0 / M Tokens Output: $6,0 / M Tokens Implicit Caching: $0,25 / M Tokens Beginnen Sie mit Qwen3.8-Max zu bauen! 🚀 📖 Blog: https://qwen.ai/blog?id=qwen3.8 ✅ Qwen Studio: https://chat.qwen.ai/?models=qwen3.8-max ⚡ API: https://www.qwencloud.com/models/qwen3.8-max
mehr auf Arint.info
#AIModel #Coding #MachineLearning #MultimodalAI #OpenWeights #Qwen3 #arint_info
-
RT @ComfyUI: MiniMax H3 ist jetzt über Partner Nodes in ComfyUI verfügbar. → Multimodale Ein- und Ausgabe: T2V, Erster/Letzter Frame, Omni Reference → Native Stereo-Audio auf jedem Clip → Bis zu 2K, 5-15 Sekunden bei 24 FPS → Charaktere, Szenen, Dialoge und Stimme können direkt vor Ort bearbeitet werden. API-Zugriff ist ab heute verfügbar. Native Open-Weight-Unterstützung folgt in Kürze. Video
mehr auf Arint.info
#APIAccess #ComfyUI #MiniMaxH3 #MultimodalAI #OpenWeight #VideoGeneration #arint_info
-
RT @ComfyUI: MiniMax H3 ist jetzt über Partner Nodes in ComfyUI verfügbar. → Multimodale Ein- und Ausgabe: T2V, Erster/Letzter Frame, Omni Reference → Native Stereo-Audio auf jedem Clip → Bis zu 2K, 5-15 Sekunden bei 24 FPS → Charaktere, Szenen, Dialoge und Stimme können direkt vor Ort bearbeitet werden. API-Zugriff ist ab heute verfügbar. Native Open-Weight-Unterstützung folgt in Kürze. Video
mehr auf Arint.info
#APIAccess #ComfyUI #MiniMaxH3 #MultimodalAI #OpenWeight #VideoGeneration #arint_info
-
RT @ComfyUI: MiniMax H3 ist jetzt über Partner Nodes in ComfyUI verfügbar. → Multimodale Ein- und Ausgabe: T2V, Erster/Letzter Frame, Omni Reference → Native Stereo-Audio auf jedem Clip → Bis zu 2K, 5-15 Sekunden bei 24 FPS → Charaktere, Szenen, Dialoge und Stimme können direkt vor Ort bearbeitet werden. API-Zugriff ist ab heute verfügbar. Native Open-Weight-Unterstützung folgt in Kürze. Video
mehr auf Arint.info
#APIAccess #ComfyUI #MiniMaxH3 #MultimodalAI #OpenWeight #VideoGeneration #arint_info
-
RT @ComfyUI: MiniMax H3 ist jetzt über Partner Nodes in ComfyUI verfügbar. → Multimodale Ein-/Ausgabe: Text-zu-Video, Erster/Letzter Frame, Omni-Referenz → Native Stereo-Audio auf jedem Clip → Bis zu 2K, 5-15 Sekunden bei 24 FPS → Charaktere, Szenen, Dialoge und Stimme können direkt vor Ort bearbeitet werden. API-Zugriff ist ab heute verfügbar. Native Open-Weight-Unterstützung folgt in Kürze. Video
mehr auf Arint.info
#APIAccess #ComfyUI #MiniMaxH3 #MultimodalAI #OpenWeight #VideoGeneration #arint_info
-
RT @ComfyUI: MiniMax H3 ist jetzt über Partner Nodes in ComfyUI verfügbar. → Multimodale Ein-/Ausgabe: Text-zu-Video, Erster/Letzter Frame, Omni-Referenz → Native Stereo-Audio auf jedem Clip → Bis zu 2K, 5-15 Sekunden bei 24 FPS → Charaktere, Szenen, Dialoge und Stimme können direkt vor Ort bearbeitet werden. API-Zugriff ist ab heute verfügbar. Native Open-Weight-Unterstützung folgt in Kürze. Video
mehr auf Arint.info
#APIAccess #ComfyUI #MiniMaxH3 #MultimodalAI #OpenWeight #VideoGeneration #arint_info
-
RT @ComfyUI: MiniMax H3 ist jetzt über Partner Nodes in ComfyUI verfügbar. → Multimodale Ein-/Ausgabe: Text-zu-Video, Erster/Letzter Frame, Omni-Referenz → Native Stereo-Audio auf jedem Clip → Bis zu 2K, 5-15 Sekunden bei 24 FPS → Charaktere, Szenen, Dialoge und Stimme können direkt vor Ort bearbeitet werden. API-Zugriff ist ab heute verfügbar. Native Open-Weight-Unterstützung folgt in Kürze. Video
mehr auf Arint.info
#APIAccess #ComfyUI #MiniMaxH3 #MultimodalAI #OpenWeight #VideoGeneration #arint_info
-
RT @MiniMax_AI: MiniMax H3: Omni-Referenz, kommerzielle Generierungsqualität, unschlagbare Kosteneffizienz, offene Gewichte. MiniMax H3: Ein offenes Modell, das die Grenzen zwischen Aufgaben und Modalitäten überwindet. Heute starten wir MiniMax H3, ein multimodales Generierungsmodell für allgemeine Zwecke. H3 versteht den kontextuellen Zusammenhang von Text, Bildern, Video und Audio und generiert Videos mit nativem Stereo-Sound, bis zu
mehr auf Arint.info
#GenerativeAI #Kosteneffizienz #MiniMaxH3 #MultimodalAI #OpenSource #VideoGeneration #arint_info
-
RT @MiniMax_AI: MiniMax H3: Omni-Referenz, kommerzielle Generierungsqualität, unschlagbare Kosteneffizienz, offene Gewichte. MiniMax H3: Ein offenes Modell, das die Grenzen zwischen Aufgaben und Modalitäten überwindet. Heute starten wir MiniMax H3, ein multimodales Generierungsmodell für allgemeine Zwecke. H3 versteht den kontextuellen Zusammenhang von Text, Bildern, Video und Audio und generiert Videos mit nativem Stereo-Sound, bis zu
mehr auf Arint.info
#GenerativeAI #Kosteneffizienz #MiniMaxH3 #MultimodalAI #OpenSource #VideoGeneration #arint_info
-
RT @MiniMax_AI: MiniMax H3: Omni-Referenz, kommerzielle Generierungsqualität, unschlagbare Kosteneffizienz, offene Gewichte. MiniMax H3: Ein offenes Modell, das die Grenzen zwischen Aufgaben und Modalitäten überwindet. Heute starten wir MiniMax H3, ein multimodales Generierungsmodell für allgemeine Zwecke. H3 versteht den kontextuellen Zusammenhang von Text, Bildern, Video und Audio und generiert Videos mit nativem Stereo-Sound, bis zu
mehr auf Arint.info
#GenerativeAI #Kosteneffizienz #MiniMaxH3 #MultimodalAI #OpenSource #VideoGeneration #arint_info
-
RT @MiniMax_AI: MiniMax H3: Omni-Referenz, kommerzielle Generierungsqualität, unschlagbare Kosteneffizienz, offene Gewichte. MiniMax H3: Ein offenes Modell, das die Grenzen zwischen Aufgaben und Modalitäten überwindet. Heute starten wir MiniMax H3, ein multimodales Generierungsmodell für allgemeine Zwecke. H3 versteht den kontextuellen Zusammenhang von Text, Bildern, Video und Audio und generiert Videos mit nativem Stereo-Sound, bis zu
mehr auf Arint.info
#GenerativeAI #Kosteneffizienz #MiniMaxH3 #MultimodalAI #OpenSource #VideoGeneration #arint_info
-
RT @MiniMax_AI: MiniMax H3: Omni-Referenz, kommerzielle Generierungsqualität, unschlagbare Kosteneffizienz, offene Gewichte. MiniMax H3: Ein offenes Modell, das die Grenzen zwischen Aufgaben und Modalitäten überwindet. Heute starten wir MiniMax H3, ein multimodales Generierungsmodell für allgemeine Zwecke. H3 versteht den kontextuellen Zusammenhang von Text, Bildern, Video und Audio und generiert Videos mit nativem Stereo-Sound, bis zu
mehr auf Arint.info
#GenerativeAI #Kosteneffizienz #MiniMaxH3 #MultimodalAI #OpenSource #VideoGeneration #arint_info
-
Black Forest Labs has launched FLUX 3 with support for 20-second video generation with synchronized audio.
#AI #FLUX #BlackForestLabs #AIVideoGeneration #MultimodalAI #GenAI #AIModels #AIImageGeneration #TextToVideo
-
Black Forest Labs has launched FLUX 3 with support for 20-second video generation with synchronized audio.
#AI #FLUX #BlackForestLabs #AIVideoGeneration #MultimodalAI #GenAI #AIModels #AIImageGeneration #TextToVideo
-
Black Forest Labs has launched FLUX 3 with support for 20-second video generation with synchronized audio.
#AI #FLUX #BlackForestLabs #AIVideoGeneration #MultimodalAI #GenAI #AIModels #AIImageGeneration #TextToVideo
-
Black Forest Labs has launched FLUX 3 with support for 20-second video generation with synchronized audio.
#AI #FLUX #BlackForestLabs #AIVideoGeneration #MultimodalAI #GenAI #AIModels #AIImageGeneration #TextToVideo
-
Black Forest Labs has launched FLUX 3 with support for 20-second video generation with synchronized audio.
#AI #FLUX #BlackForestLabs #AIVideoGeneration #MultimodalAI #GenAI #AIModels #AIImageGeneration #TextToVideo
-
Kobe Steel invests in Noetra for Japan multimodal AI project
KEY POINTSKobe Steel invests in Noetra to join development of a domestic Japanese multimodal foundation modelProject supports physical…
#EuropeSays #Japan #JP #Japanese #KobeSteel #Kobelco #METI #multimodalAI #NEDO #Noetra #physicalai
https://www.europesays.com/japan/62415/ -
https://winbuzzer.com/2026/07/20/google-vids-rolls-out-personal-avatars-and-gemini-omni-xcxwbn/
Google Vids has started rolling out Personal AI Avatar creation and Gemini Omni for eligible customers, with account-bound likenesses & invisible SynthID watermarks.
#AI #GoogleVids #GeminiOmni #AIAvatars #Google #GoogleGemini #GoogleAI #GenAI #MultimodalAI #AIVideo
-
https://winbuzzer.com/2026/07/20/google-vids-rolls-out-personal-avatars-and-gemini-omni-xcxwbn/
Google Vids has started rolling out Personal AI Avatar creation and Gemini Omni for eligible customers, with account-bound likenesses & invisible SynthID watermarks.
#AI #GoogleVids #GeminiOmni #AIAvatars #Google #GoogleGemini #GoogleAI #GenAI #MultimodalAI #AIVideo
-
https://winbuzzer.com/2026/07/20/google-vids-rolls-out-personal-avatars-and-gemini-omni-xcxwbn/
Google Vids has started rolling out Personal AI Avatar creation and Gemini Omni for eligible customers, with account-bound likenesses & invisible SynthID watermarks.
#AI #GoogleVids #GeminiOmni #AIAvatars #Google #GoogleGemini #GoogleAI #GenAI #MultimodalAI #AIVideo
-
https://winbuzzer.com/2026/07/20/google-vids-rolls-out-personal-avatars-and-gemini-omni-xcxwbn/
Google Vids has started rolling out Personal AI Avatar creation and Gemini Omni for eligible customers, with account-bound likenesses & invisible SynthID watermarks.
#AI #GoogleVids #GeminiOmni #AIAvatars #Google #GoogleGemini #GoogleAI #GenAI #MultimodalAI #AIVideo
-
https://winbuzzer.com/2026/07/20/google-vids-rolls-out-personal-avatars-and-gemini-omni-xcxwbn/
Google Vids has started rolling out Personal AI Avatar creation and Gemini Omni for eligible customers, with account-bound likenesses & invisible SynthID watermarks.
#AI #GoogleVids #GeminiOmni #AIAvatars #Google #GoogleGemini #GoogleAI #GenAI #MultimodalAI #AIVideo
-
https://winbuzzer.com/2026/07/17/moonshot-ai-unveils-28t-parameter-kimi-k3-ai-model-xcxwbn/
Moonshot AI has launched its 2.8-trillion-parameter Kimi K3 model, but a high hallucination rate might tempers its frontier-model pitch.
#AI #MoonshotAI #KimiK3 #AIModels #MultimodalAI #MixtureOfExperts #AIReasoningModels #AIBenchmarks #ChinaAI
-
https://winbuzzer.com/2026/07/17/moonshot-ai-unveils-28t-parameter-kimi-k3-ai-model-xcxwbn/
Moonshot AI has launched its 2.8-trillion-parameter Kimi K3 model, but a high hallucination rate might tempers its frontier-model pitch.
#AI #MoonshotAI #KimiK3 #AIModels #MultimodalAI #MixtureOfExperts #AIReasoningModels #AIBenchmarks #ChinaAI
-
https://winbuzzer.com/2026/07/17/moonshot-ai-unveils-28t-parameter-kimi-k3-ai-model-xcxwbn/
Moonshot AI has launched its 2.8-trillion-parameter Kimi K3 model, but a high hallucination rate might tempers its frontier-model pitch.
#AI #MoonshotAI #KimiK3 #AIModels #MultimodalAI #MixtureOfExperts #AIReasoningModels #AIBenchmarks #ChinaAI
-
https://winbuzzer.com/2026/07/17/moonshot-ai-unveils-28t-parameter-kimi-k3-ai-model-xcxwbn/
Moonshot AI has launched its 2.8-trillion-parameter Kimi K3 model, but a high hallucination rate might tempers its frontier-model pitch.
#AI #MoonshotAI #KimiK3 #AIModels #MultimodalAI #MixtureOfExperts #AIReasoningModels #AIBenchmarks #ChinaAI
-
https://winbuzzer.com/2026/07/17/moonshot-ai-unveils-28t-parameter-kimi-k3-ai-model-xcxwbn/
Moonshot AI has launched its 2.8-trillion-parameter Kimi K3 model, but a high hallucination rate might tempers its frontier-model pitch.
#AI #MoonshotAI #KimiK3 #AIModels #MultimodalAI #MixtureOfExperts #AIReasoningModels #AIBenchmarks #ChinaAI
-
https://www.europesays.com/ie/572647/ Jason Calacanis Says Nvidia Is ‘Taking the Gloves Off’ With Nemotron, Predicts Jensen Huang Will Challenge OpenAI, Anthropic by Owning the Whole AI Stack #AI #ArtificialIntelligence #ArtificialIntelligence #Éire #IE #Ireland #JasonCalacanis #JensenHuang #MultimodalAI #Nvidia #OpenSourceModel #OpenAI #Technology
-
https://www.europesays.com/people/142412/ Jason Calacanis Says Nvidia Is ‘Taking the Gloves Off’ With Nemotron, Predicts Jensen Huang Will Challenge OpenAI, Anthropic by Owning the Whole AI Stack #JasonCalacanis #JensenHuang #MultimodalAI #Nvidia #OpenSourceModel #OpenAI
-
Samsung Launches Galaxy XR in the UK https://www.byteseu.com/2115371/ #AndroidEnterprise #AndroidXR #Enterprise #ExtendedReality #GalaxyXR #GreatBritain #MultimodalAI #UnitedKingdom #Wearable
-
NAVER D2SF Invests in AIM Intelligence, an AI Security Startup
– Rising AI adoption drives growing security risks, making AI security essential infrastructure– End-to-end AI security solutions spanning…
#EuropeSays #Korea #KR #Naver #AIsecurityandsafety #generativeAI #intelligence #multimodalAI #SangyoonYu #securityrisks #securitysolutions
https://www.europesays.com/korea/49725/ -
Google redesigned Workspace icons for people. Embedding analysis suggests the new designs are also easier for vision models to distinguish. https://hackernoon.com/did-googles-workspace-redesign-make-its-icons-easier-for-ai-to-see #multimodalai
-
Google redesigned Workspace icons for people. Embedding analysis suggests the new designs are also easier for vision models to distinguish. https://hackernoon.com/did-googles-workspace-redesign-make-its-icons-easier-for-ai-to-see #multimodalai
-
Google redesigned Workspace icons for people. Embedding analysis suggests the new designs are also easier for vision models to distinguish. https://hackernoon.com/did-googles-workspace-redesign-make-its-icons-easier-for-ai-to-see #multimodalai
-
https://winbuzzer.com/2026/06/06/alibaba-pitches-qwen37-plus-as-a-computer-use-ai-agent-xcxwbn/
Alibaba's new Qwen3.7-Plus model targets screen, coding, and cloud-console automation as computer-use AI pushes beyond browsers into app and terminal tasks.
#AI #Qwen37Plus #Qwen3 #Alibaba #Qwen #AIAutomation #AIAgents #MultimodalAI #AICoding #AIModels
-
https://winbuzzer.com/2026/06/06/alibaba-pitches-qwen37-plus-as-a-computer-use-ai-agent-xcxwbn/
Alibaba's new Qwen3.7-Plus model targets screen, coding, and cloud-console automation as computer-use AI pushes beyond browsers into app and terminal tasks.
#AI #Qwen37Plus #Qwen3 #Alibaba #Qwen #AIAutomation #AIAgents #MultimodalAI #AICoding #AIModels
-
https://winbuzzer.com/2026/06/04/google-gemma-4-12b-targets-local-ai-agents-on-laptops-xcxwbn/
Google has released Gemma 4 12B, a local multimodal AI model for laptops that handles audio, images, code, and tool calls with 16GB memory locally.
#AI #Gemma4E4B #Gemma4 #GoogleGemma #Gemma #Google #GoogleAI #GoogleDeepMind #AIModels #MultimodalAI #OpenSourceAI #OnDeviceAI #AIAgents #AgenticAI
-
Google wants your next AI agent running locally on a 16GB laptop
https://web.brid.gy/r/https://nerds.xyz/2026/06/google-gemma-4-12b-local-ai/
-
https://winbuzzer.com/2026/06/02/microsoft-adds-seven-mai-models-to-foundry-for-developers-xcxwbn/
Microsoft is putting seven first-party MAI models into developer channels, led by the MAI-Thinking-1 reasoning model in Foundry private preview.
#AI #MAIThinking1 #MicrosoftFoundry #Microsoft #MicrosoftAI #AIModels #MultimodalAI #Build2026