الأخبار التي صنعها الإنسان مقابل الأخبار التي أنشأها الذكاء الاصطناعي: مقارنة تقييمات طلاب الصحافة البرتغاليين والإسبان
Human-made news vs AI-generated news: a comparison of Portuguese and Spanish journalism students’ evaluations

شارك:
المجلة: Humanities and Social Sciences Communications، المجلد: 12، العدد: 1
DOI: https://doi.org/10.1057/s41599-025-04872-2
تاريخ النشر: 2025-04-24
المؤلف: João Pedro Baptista وآخرون
الموضوع الرئيسي: التواصل وتأثير COVID-19

نظرة عامة

تستكشف هذه الدراسة تصورات طلاب الجامعات في البرتغال وإسبانيا بشأن جودة الأخبار الصحفية التي تنتجها الذكاء الاصطناعي التوليدي (AI)، مع التركيز بشكل خاص على نماذج مثل ChatGPT-3. مع تزايد انتشار الذكاء الاصطناعي في توزيع الأخبار ووظائف غرف الأخبار، يعد بتحسين الكفاءة ولكنه يثير أيضًا مخاوف أخلاقية وأسئلة حول العلاقة بين الصحفيين البشريين والمحتوى الذي ينتجه الذكاء الاصطناعي. تتناول الأبحاث أسئلة رئيسية حول قدرة الطلاب على التمييز بين الأخبار التي ينتجها الذكاء الاصطناعي وتلك التي يكتبها الصحفيون، وجودة هذه الأنواع من الأخبار، وتأثير مواضيع الأخبار على تقييمات الجودة.

باستخدام استطلاعين تم إجراؤهما على 444 مشاركًا، قيمت الدراسة أبعادًا مختلفة من جودة الأخبار، بما في ذلك قابلية القراءة، والمعلوماتية، والأسلوب. أظهرت النتائج أن الطلاب عمومًا قيموا الأخبار التي ينتجها الذكاء الاصطناعي من ChatGPT-3 على أنها ذات جودة أعلى من تلك التي ينتجها الصحفيون البشريون، حيث قدم الطلاب الإسبان تقييمات إيجابية بشكل خاص للمحتوى الذي ينتجه الذكاء الاصطناعي. تشير هذه النتائج إلى تحول محتمل في تصورات جودة الأخبار، مما يبرز الحاجة إلى مناقشات مستمرة حول آثار الذكاء الاصطناعي في تعليم الصحافة والمشهد الإعلامي الأوسع.

مقدمة

تسلط مقدمة هذه الورقة البحثية الضوء على التأثير التحويلي للذكاء الاصطناعي (AI) على مشهد الصحافة، لا سيما في توزيع محتوى الأخبار وعمليات غرف الأخبار. لقد غير الذكاء الاصطناعي، وخاصة أدوات الذكاء الاصطناعي التوليدية مثل ChatGPT، الأدوار التقليدية للصحفيين، مما مكن من الأتمتة في إنشاء الأخبار وأثار مخاوف بشأن جودة ومصداقية المعلومات. بينما تشير بعض الدراسات إلى أن الذكاء الاصطناعي يمكن أن يعزز كفاءة الصحافة من خلال تخفيف المهام المتكررة، هناك مخاوف كبيرة بشأن انتشار المعلومات المضللة والمشكلات الأخلاقية المرتبطة بالمحتوى الذي ينتجه الذكاء الاصطناعي. تشير الورقة إلى وجود ثنائية في الأدبيات، حيث يدعو بعض العلماء إلى الذكاء الاصطناعي كأداة مفيدة للصحافة، بينما يعبر آخرون عن شكوكهم بشأن آثارها على الصناعة.

تتناول الأبحاث بشكل خاص حالة تعليم الذكاء الاصطناعي في برامج الصحافة في إسبانيا والبرتغال، كاشفة عن نقص في التدريب الشامل على تقنيات الذكاء الاصطناعي بين طلاب الصحافة. على الرغم من وجود أدلة على تزايد وجود الذكاء الاصطناعي في غرف الأخبار، لا يزال هناك فجوة في فهم واستخدام هذه التقنيات بشكل فعال. تهدف الدراسة إلى استكشاف تصورات طلاب الصحافة حول جودة الأخبار التي ينتجها الذكاء الاصطناعي وقدرتهم على تمييزها عن المحتوى الذي يكتبه البشر، مما يساهم في النقاش المستمر حول دمج الذكاء الاصطناعي في تعليم وممارسة الصحافة. تركز الأسئلة البحثية المطروحة على تمييز الطلاب بين الأخبار التي ينتجها الذكاء الاصطناعي وتلك التي يكتبها البشر وتأثير موضوعات الأخبار على الجودة المدركة.

الطرق

تستكشف هذه الدراسة تصورات طلاب الجامعات في البرتغال وإسبانيا بشأن جودة النصوص للأخبار التي تم إنشاؤها بواسطة ChatGPT-3 مقارنة بتلك التي كتبها صحفيون محترفون. أجريت الدراسة خلال الفصل الدراسي الأول من العام الأكاديمي 2023-2024، وشملت استطلاعين منفصلين في الفصول الدراسية، مما أسفر عن عينة إجمالية من 444 طالبًا – 99 من البرتغال و345 من إسبانيا. شملت المشاركين، الذين تتراوح أعمارهم بين 18 و65 عامًا، 57.9% إناث و41.9% ذكور، مع معدل استجابة 100% بسبب الإدارة داخل الفصل للاستطلاعات. التركيز على طلاب الصحافة مهم، حيث يمكن أن تؤثر آرائهم حول الذكاء الاصطناعي التوليدي على الممارسات الصحفية المستقبلية، لا سيما فيما يتعلق بالأخلاقيات وجودة الأخبار.

تضمنت المنهجية ثلاث إجراءات رئيسية: أولاً، اختيار وجمع مقالات الأخبار التي كتبها الصحفيون؛ ثانيًا، استخراج المحفزات من هذه المقالات لـ ChatGPT-3 لإنشاء محتوى أخبار مكافئ، باستخدام ستة كلمات رئيسية بناءً على أسئلة Lead الصحفية التقليدية (الستة W’s: من، ماذا، متى، أين، لماذا، كيف)؛ وثالثًا، تطوير استطلاعات باللغتين البرتغالية والإسبانية لتناسب السياق اللغوي للمشاركين. يستند اختيار إسبانيا والبرتغال إلى قربهما الثقافي وأنظمة الإعلام المشتركة، على الرغم من الاختلافات الملحوظة في الاستقطاب السياسي، والتي قد تؤثر على تقييمات الطلاب للأخبار التي ينتجها الذكاء الاصطناعي.

النتائج

تشير النتائج إلى أن طلاب الجامعات الإيبريين يرون أن الأخبار التي ينتجها ChatGPT ذات جودة أعلى من تلك التي ينتجها الصحفيون، مع تقييمات متوسطة قدرها $M = 4.85$ (SD = 0.82) لـ ChatGPT و$M = 4.30$ (SD = 0.84) للصحفيين. هذه الفجوة ذات دلالة إحصائية ($t(444) = 11.919, p < 0.001, d = 0.566$). من الجدير بالذكر أن الطلاب الإسبان قيموا أخبار ChatGPT أعلى ($M = 4.92$, SD = 0.80) مقارنة بالطلاب البرتغاليين ($M = 4.59$, SD = 0.84). فيما يتعلق بالأخبار التي ينتجها الصحفيون، قدم الطلاب الإسبان أدنى تقييم متوسط ($M = 4.33$, SD = 0.79). أكدت تحليل التباين الأحادي (ANOVA) وجود اختلافات كبيرة في الجودة المدركة لأخبار ChatGPT بين البلدين ($F(1,444) = 12.238, p < 0.001$). عبر فئات الأخبار المختلفة، حصلت أخبار ChatGPT باستمرار على تقييمات أعلى، لا سيما في الثقافة ($M = 4.92$, SD = 1.01) والحوادث ($M = 4.76$, SD = 1.09) في البرتغال. على العكس من ذلك، تم تقييم أخبار الجرائم على أنها الأدنى بشكل عام. كشفت مؤشرات الجودة أن الصحفيين واجهوا صعوبة في تقديم المصادر وعلامات الترقيم، حيث سجلوا $M = 4.03$ (SD = 1.11) للترقيم، بينما سجل ChatGPT درجات أعلى بكثير ($M = 4.73$, SD = 0.99$; $t(444) = -10.899, p < 0.001, d = -0.517$). بشكل عام، تفوق ChatGPT على الصحفيين في قابلية القراءة، وعرض الحقائق، ووضوح اللغة، مع ملاحظات كبيرة في تقييمات قابلية القراءة وجودة اللغة بين الطلاب الإسبان والبرتغاليين. ومع ذلك، لم تظهر معظم المؤشرات اختلافات كبيرة بناءً على البلد، مما يشير إلى توافق عام حول الجودة المدركة للأخبار التي ينتجها ChatGPT.

المناقشة

فحصت الدراسة الجودة المدركة للأخبار التي ينتجها ChatGPT-3 مقارنة بتلك التي ينتجها الصحفيون البشريون بين طلاب الجامعات في البرتغال وإسبانيا. أشارت النتائج إلى تفضيل كبير للأخبار التي ينتجها الذكاء الاصطناعي، والتي حصلت على تقييمات أعلى من حيث قابلية القراءة، والمعلوماتية، وبنية النص عبر فئات مختلفة، بما في ذلك الحوادث، والجرائم، والمشاهير، والثقافة. من الجدير بالذكر أن الطلاب الإسبان قيموا الأخبار التي ينتجها الذكاء الاصطناعي بشكل أكثر إيجابية من نظرائهم البرتغاليين، مما يشير إلى أن أسلوب كتابة ChatGPT يتماشى بشكل أقرب مع المعايير التعليمية في إسبانيا. يثير هذا التفضيل تساؤلات حول الجودة الحالية للصحافة البشرية، حيث قيم الطلاب الأخبار التي أنشأها الصحفيون بشكل أقل، لا سيما فيما يتعلق بعرض المصادر والمراجع.

تشير النتائج إلى احتمال تراجع جودة النصوص للأخبار التي كتبها البشر، ربما تأثراً بعوامل مثل الضغط للنشر والطبيعة الاستفزازية للصحافة المعاصرة. تسلط الدراسة الضوء على القدرات المتزايدة للذكاء الاصطناعي في إنشاء الأخبار، مما يتحدى الرؤية التقليدية للصحفيين البشر كمعيار ذهبي. بينما النتائج محددة لطلاب الجامعات، فإنها تثير اعتبارات أوسع حول آثار الذكاء الاصطناعي على الصحافة والمشهد المتطور لاستهلاك الأخبار. يجب أن تستكشف الأبحاث المستقبلية هذه الديناميكيات بشكل أعمق، لا سيما فيما يتعلق بالمعايير والممارسات التحريرية التي قد تؤثر على الجودة المدركة للأخبار.

Journal: Humanities and Social Sciences Communications, Volume: 12, Issue: 1
DOI: https://doi.org/10.1057/s41599-025-04872-2
Publication Date: 2025-04-24
Author(s): João Pedro Baptista et al.
Primary Topic: Communication and COVID-19 Impact

Overview

This study investigates the perceptions of university students in Portugal and Spain regarding the quality of journalistic news produced by generative artificial intelligence (AI), specifically focusing on models like ChatGPT-3. As AI becomes more prevalent in news distribution and newsroom functions, it promises increased efficiency but also raises ethical concerns and questions about the relationship between human journalists and AI-generated content. The research addresses key questions about students’ abilities to differentiate between AI-generated and journalist-written news, the comparative quality of these news types, and the impact of news topics on quality assessments.

Utilizing two surveys administered to 444 participants, the study evaluated various dimensions of news quality, including readability, informativeness, and style. Results revealed that students generally rated AI-generated news from ChatGPT-3 as higher quality than that produced by human journalists, with Spanish students providing particularly favorable evaluations of AI-generated content. These findings indicate a potential shift in perceptions of news quality, highlighting the need for ongoing discussions about the implications of AI in journalism education and the broader media landscape.

Introduction

The introduction of this research paper highlights the transformative impact of artificial intelligence (AI) on the journalism landscape, particularly in news content distribution and newsroom operations. AI, particularly generative AI tools like ChatGPT, has shifted the traditional roles of journalists, enabling automation in news creation and raising concerns about the quality and credibility of information. While some studies suggest that AI can enhance journalistic efficiency by alleviating repetitive tasks, there are significant apprehensions regarding the proliferation of disinformation and ethical dilemmas associated with AI-generated content. The paper notes a dichotomy in the literature, with some scholars advocating for AI as a beneficial tool for journalism, while others express skepticism about its implications for the industry.

The research specifically addresses the state of AI education in journalism programs in Spain and Portugal, revealing a lack of comprehensive training on AI technologies among journalism students. Despite evidence of AI’s growing presence in newsrooms, particularly in Spain, there remains a gap in understanding and utilizing these technologies effectively. The study aims to explore journalism students’ perceptions of AI-generated news quality and their ability to distinguish it from human-written content, thereby contributing to the ongoing discourse on the integration of AI in journalism education and practice. The research questions posed focus on students’ discernment of AI-generated versus human-written news and the influence of news subject matter on perceived quality.

Methods

This study investigates the perceptions of university students in Portugal and Spain regarding the textual quality of news articles generated by ChatGPT-3 compared to those written by professional journalists. Conducted during the first semester of the 2023-2024 academic year, the research involved two separate classroom surveys, yielding a total sample of 444 students—99 from Portugal and 345 from Spain. The participants, aged between 18 and 65 years, included 57.9% females and 41.9% males, with a 100% response rate due to the in-class administration of the surveys. The focus on journalism students is significant, as their views on generative AI could influence future journalistic practices, particularly concerning ethics and news quality.

The methodology comprised three key procedures: first, the selection and collection of news articles authored by journalists; second, the extraction of prompts from these articles for ChatGPT-3 to generate equivalent news content, utilizing six keywords based on the traditional journalistic Lead questions (the six W’s: who, what, when, where, why, how); and third, the development of surveys in both Portuguese and Spanish to accommodate the linguistic context of the participants. The choice of Spain and Portugal is grounded in their cultural proximity and shared media systems, despite notable differences in political polarization, which may affect students’ evaluations of AI-generated news.

Results

The results indicate that Iberian university students perceive news generated by ChatGPT to be of higher quality than that produced by journalists, with average ratings of $M = 4.85$ (SD = 0.82) for ChatGPT and $M = 4.30$ (SD = 0.84) for journalists. This difference is statistically significant ($t(444) = 11.919, p < 0.001, d = 0.566$). Notably, Spanish students rated ChatGPT news higher ($M = 4.92$, SD = 0.80) compared to Portuguese students ($M = 4.59$, SD = 0.84). In terms of journalist-produced news, Spanish students provided the lowest average rating ($M = 4.33$, SD = 0.79). Univariate analysis of variance (ANOVA) confirmed significant differences in perceived quality of ChatGPT news between the two countries ($F(1,444) = 12.238, p < 0.001$). Across various news categories, ChatGPT news consistently received higher ratings, particularly in culture ($M = 4.92$, SD = 1.01) and accidents ($M = 4.76$, SD = 1.09) in Portugal. Conversely, crime news was rated the lowest overall. Quality indicators revealed that journalists struggled with presenting sources and punctuation, scoring $M = 4.03$ (SD = 1.11) for punctuation, while ChatGPT scored significantly higher ($M = 4.73$, SD = 0.99$; $t(444) = -10.899, p < 0.001, d = -0.517$). Overall, ChatGPT outperformed journalists in readability, presentation of facts, and clarity of language, with significant differences noted in readability and language quality ratings between Spanish and Portuguese students. However, most indicators did not show significant differences based on country, suggesting a general consensus on the perceived quality of news generated by ChatGPT.

Discussion

The study examined the perceived quality of news generated by ChatGPT-3 compared to that produced by human journalists among university students in Portugal and Spain. Results indicated a significant preference for AI-generated news, which received higher ratings in terms of readability, informativeness, and textual structure across various categories, including accidents, crime, celebrity, and culture. Notably, Spanish students rated the AI-generated news even more favorably than their Portuguese counterparts, suggesting that the writing style of ChatGPT aligns more closely with the educational standards in Spain. This preference raises questions about the current quality of human journalism, as students rated journalist-created news lower, particularly regarding the presentation of sources and references.

The findings suggest a potential decline in the textual quality of human-written news, possibly influenced by factors such as the pressure to publish and the sensationalist nature of contemporary journalism. The study highlights the growing capabilities of AI in news generation, challenging the traditional view of human journalists as the gold standard. While the results are specific to university students, they prompt broader considerations about the implications of AI on journalism and the evolving landscape of news consumption. Future research should explore these dynamics further, particularly in relation to the editorial standards and practices that may affect the perceived quality of news.

شارك: