Aphantasia Logo
Back to all research
Aphantasia Logo

Building awareness and understanding of aphantasia through research, education, and community support.

About

  • What is Aphantasia?
  • What is Hyperphantasia?
  • Take Assessment
  • Getting Started
  • Newsletter
  • About Us
  • Contact

Community

  • Premium Membership
  • Find support
  • Discussions
  • Events
  • Visualize

For Professionals

  • Overview
  • Free Introduction
  • Counselor Training
  • Educator Training
  • List Your Practice
  • Pricing & Bundles

Resources

  • Articles & Stories
  • Videos & Interviews
  • Aphantasia Course
  • FAQs

Research

  • Research Library
  • Participate in Studies
  • Recruitment Services

© 2026 Aphantasia Network. All rights reserved.

  • Terms and Conditions
  • Privacy Policy
Aphantasia Logo
Back to all research
Aphantasia Logo

Building awareness and understanding of aphantasia through research, education, and community support.

About

  • What is Aphantasia?
  • What is Hyperphantasia?
  • Take Assessment
  • Getting Started
  • Newsletter
  • About Us
  • Contact

Community

  • Premium Membership
  • Find support
  • Discussions
  • Events
  • Visualize

For Professionals

  • Overview
  • Free Introduction
  • Counselor Training
  • Educator Training
  • List Your Practice
  • Pricing & Bundles

Resources

  • Articles & Stories
  • Videos & Interviews
  • Aphantasia Course
  • FAQs

Research

  • Research Library
  • Participate in Studies
  • Recruitment Services

© 2026 Aphantasia Network. All rights reserved.

  • Terms and Conditions
  • Privacy Policy
Ask AI About This Paper
Ask AI About This Paper

Reliability generalization meta-analysis of the internal consistency of the Big Five Inventory (BFI) by comparing BFI (44 items) and BFI-2 (60 items) versions controlling for age, sex, language factors

DOI: 10.1186/s40359-024-02271-x
Tags:
Measurement & Assessment
Husain, W., Haddad, A. J., Husain, M. A., Ghazzawi, H., Trabelsi, K., Ammar, A., Saif, Z., Pakpour, A., & Jahrami, H. (2025). Reliability generalization meta-analysis of the internal consistency of the big five inventory (bfi) by comparing bfi (44 items) and bfi-2 (60 items) versions controlling for age, sex, language factors. BMC Psychology, 13(1). doi:/10.1186/s40359-024-02271-x

Abstract

The Big Five Inventory (BFI) is a popular measure that evaluates personality on the Big-Five model. Apart from its utilization across cultures, the literature did not reveal any meta-analysis for the reliability of the different versions of the BFI and its translations. The current study carried out a reliability generalization meta-analysis (REGEMA) to establish the reliability of the BFI across cultures and languages. We searched 30 databases for the relevant studies from 1991 to mid-November 2024. The studies that we intended to include in our meta-analysis required to have utilized the BFI (44 items) and the BFI-2 (60 items) and have reported Cronbach's alpha or McDonald's omega reliability estimates. Our coded variables included BFI version, sample size, population type, age, gender, clinical state, and reliability. A total of 57 studies (datapoints) published in 34 research articles (involving 43,715 participants; 60.24% women; Mean age = 30.08) from various cultures and languages were finally included. These studies used BFI and BFI-2 in Arabic, Chinese, Croatian, Czech, Danish, Dutch, English, French, German, Indonesian, Italian, Japanese, Malay, Norwegian, Polish, Portuguese, Russian, Serbian, Spanish, Swahili, and Turkish. Data analysis was conducted using the metafor and meta packages in R. The average correlation was computed using a random-effects model and reliability coefficients indicated effect size. I2 and Cochran's Q tests were used to examine heterogeneity, with prediction intervals suggesting genuine influences around the pooled estimate. Using funnel plots, regression-based tests (e.g., Egger's regression, rank correlation), and trim-and-fill imputation, publication bias was adjusted to estimate unbiased effects. We calculated the individual and combined reliability of the BFI and BFI-2 across languages and cultures. The results revealed the reliability of all five factors used in the BFI/BFI-2. The BFI estimates provide the following results: openness is estimated at 0.77 (95% CI: 0.75; 0.80); conscientiousness is estimated at 0.80 (95% CI: 0.78; 0.82); extraversion is also estimated at 0.80 (95% CI: 0.79; 0.82); agreeableness is estimated at 0.73 (95% CI: 0.71; 0.76); and neuroticism is estimated at 0.80 (95% CI: 0.79; 0.82). The BFI-2 estimates are as follows: openness is estimated at 0.83 (95% CI: 0.82; 0.84); conscientiousness is estimated at 0.86 (95% CI: 0.85; 0.87); extraversion is estimated at 0.85 (95% CI: 0.84; 0.86); agreeableness is also estimated at 0.80 (95% CI: 0.79; 81); and neuroticism is estimated at 0.89 (95% CI: 0.88; 0.89). The current meta-analysis represents the first reliability analysis of the BFI and the first comparison between its two different versions, the BFI (44 items) and the BFI-2 (60 items). The generalized reliability of both the BFI and BFI-2 were established. The findings confirm that the BFI and BFI-2 have good reliability across all five factors.

Authors

  • Waqar Husain1
  • Areen Jamal Haddad1
  • Muhammad Ahmad Husain1
  • Hadeel Ghazzawi1
  • Khaled Trabelsi1
  • Achraf Ammar1
  • Zahra Saif1
  • Amir Pakpour1
  • Haitham Jahrami1

Reliability generalization meta-analysis of the internal consistency of the Big Five Inventory (BFI) by comparing BFI (44 items) and BFI-2 (60 items) versions controlling for age, sex, language factors

DOI: 10.1186/s40359-024-02271-x
Tags:
Measurement & Assessment
Husain, W., Haddad, A. J., Husain, M. A., Ghazzawi, H., Trabelsi, K., Ammar, A., Saif, Z., Pakpour, A., & Jahrami, H. (2025). Reliability generalization meta-analysis of the internal consistency of the big five inventory (bfi) by comparing bfi (44 items) and bfi-2 (60 items) versions controlling for age, sex, language factors. BMC Psychology, 13(1). doi:/10.1186/s40359-024-02271-x

Abstract

The Big Five Inventory (BFI) is a popular measure that evaluates personality on the Big-Five model. Apart from its utilization across cultures, the literature did not reveal any meta-analysis for the reliability of the different versions of the BFI and its translations. The current study carried out a reliability generalization meta-analysis (REGEMA) to establish the reliability of the BFI across cultures and languages. We searched 30 databases for the relevant studies from 1991 to mid-November 2024. The studies that we intended to include in our meta-analysis required to have utilized the BFI (44 items) and the BFI-2 (60 items) and have reported Cronbach's alpha or McDonald's omega reliability estimates. Our coded variables included BFI version, sample size, population type, age, gender, clinical state, and reliability. A total of 57 studies (datapoints) published in 34 research articles (involving 43,715 participants; 60.24% women; Mean age = 30.08) from various cultures and languages were finally included. These studies used BFI and BFI-2 in Arabic, Chinese, Croatian, Czech, Danish, Dutch, English, French, German, Indonesian, Italian, Japanese, Malay, Norwegian, Polish, Portuguese, Russian, Serbian, Spanish, Swahili, and Turkish. Data analysis was conducted using the metafor and meta packages in R. The average correlation was computed using a random-effects model and reliability coefficients indicated effect size. I2 and Cochran's Q tests were used to examine heterogeneity, with prediction intervals suggesting genuine influences around the pooled estimate. Using funnel plots, regression-based tests (e.g., Egger's regression, rank correlation), and trim-and-fill imputation, publication bias was adjusted to estimate unbiased effects. We calculated the individual and combined reliability of the BFI and BFI-2 across languages and cultures. The results revealed the reliability of all five factors used in the BFI/BFI-2. The BFI estimates provide the following results: openness is estimated at 0.77 (95% CI: 0.75; 0.80); conscientiousness is estimated at 0.80 (95% CI: 0.78; 0.82); extraversion is also estimated at 0.80 (95% CI: 0.79; 0.82); agreeableness is estimated at 0.73 (95% CI: 0.71; 0.76); and neuroticism is estimated at 0.80 (95% CI: 0.79; 0.82). The BFI-2 estimates are as follows: openness is estimated at 0.83 (95% CI: 0.82; 0.84); conscientiousness is estimated at 0.86 (95% CI: 0.85; 0.87); extraversion is estimated at 0.85 (95% CI: 0.84; 0.86); agreeableness is also estimated at 0.80 (95% CI: 0.79; 81); and neuroticism is estimated at 0.89 (95% CI: 0.88; 0.89). The current meta-analysis represents the first reliability analysis of the BFI and the first comparison between its two different versions, the BFI (44 items) and the BFI-2 (60 items). The generalized reliability of both the BFI and BFI-2 were established. The findings confirm that the BFI and BFI-2 have good reliability across all five factors.

Authors

  • Waqar Husain1
  • Areen Jamal Haddad1
  • Muhammad Ahmad Husain1
  • Hadeel Ghazzawi1
  • Khaled Trabelsi1
  • Achraf Ammar1
  • Zahra Saif1
  • Amir Pakpour1
  • Haitham Jahrami1
Aphantasia Logo

What This Study Is About

This study examines how reliably researchers can measure personality across different cultures and languages. It specifically looks at the "Big Five Inventory," a popular tool used to assess five major personality traits: openness, conscientiousness, extraversion, agreeableness, and neuroticism.

How They Studied It

The researchers conducted a "meta-analysis," which means they combined and analyzed data from many previous studies. They looked at 70 different research papers that used the Big Five Inventory (BFI) or its updated version (BFI-2). In total, these studies included over 43,000 participants from a wide variety of countries, including Jordan, Germany, Pakistan, and China. The team checked how consistently the questions in these surveys measured the same personality traits across different groups.

What They Found

The study found that both versions of the personality test are generally very reliable tools for scientists to use worldwide. The newer version (BFI-2) was slightly more consistent than the original. However, the researchers noticed that the results were less consistent in certain cultures or when translated into specific languages. For example, the trait of "openness"—which includes curiosity and the use of mental imagery (the ability to picture things in your mind)—showed more variation between different countries than other traits.

What This Might Mean

These findings suggest that while the Big Five Inventory is a strong tool, researchers must be careful when comparing personality scores between different cultures. A person’s score might be influenced by how their specific culture understands certain concepts, rather than just their personality. For the aphantasia community, this is important because it shows that "openness" (a trait often linked to vivid imagination) is measured slightly differently depending on where a person lives or what language they speak.

One Interesting Detail

The researchers found that the "extraversion" trait was the most consistently measured across all the different studies and languages, while "openness" was the most difficult to measure consistently across different cultures.
This summary was generated by AI and may contain errors. Always refer to the original paper for accuracy.
Most-Cited Researchers
Top 10 researchers by number of papers published in Aphantasia
1

Joel Pearson

35
2

Merlin Monzel

30
3

Adam Zeman

20
4

Rebecca Keogh

17
5

Martin Reuter

16
6

Juha Silvanto

14
7

Paolo Bartolomeo

11
8

Carla Dance

10
9

Jianghao Liu

9
10

Fraser Milton

9
Recruit the largest aphantasia cohort on earth

97.4K+ users & subscribers across the imagery spectrum, plus 1.5M+ survey responses in our dataset. Skip the recruitment bottleneck and run your study with the community we've built for a decade.

Partner with us