Vernacular? I Barely Know Her: Challenges with Style Control and Stereotyping

Aich, Ankit, Liu, Tingting, Giorgi, Salvatore, Isman, Kelsey, Ungar, Lyle, Curtis, Brenda

Jun-18-2024–arXiv.org Artificial Intelligence

Large Language Models (LLMs) are increasingly being used in educational and learning applications. Research has demonstrated that controlling for style, to fit the needs of the learner, fosters increased understanding, promotes inclusion, and helps with knowledge distillation. To understand the capabilities and limitations of contemporary LLMs in style control, we evaluated five state-of-the-art models: GPT-3.5, GPT-4, GPT-4o, Llama-3, and Mistral-instruct-7B across two style control tasks. We observed significant inconsistencies in the first task, with model performances averaging between 5th and 8th grade reading levels for tasks intended for first-graders, and standard deviations up to 27.6. For our second task, we observed a statistically significant improvement in performance from 0.02 to 0.26. However, we find that even without stereotypes in reference texts, LLMs Figure 1: Overall view of this paper. We find that while often generated culturally insensitive content in-context learning can control for reading level and during their tasks. We provide a thorough analysis simplicity, it cannot do the same for vernacular English.

instruction, reading level, stereotype, (17 more...)

arXiv.org Artificial Intelligence

Jun-18-2024

arXiv.org PDF

Add feedback

Country:
- Africa > Ghana
  - Eastern Region > Koforidua (0.04)
- Asia > Taiwan (0.04)
- Europe > Italy
  - Tuscany > Florence (0.04)
- North America
  - Canada > Ontario
    - Toronto (0.04)
  - United States
    - California (0.04)
    - Nebraska > Lancaster County
      - Lincoln (0.04)
    - Pennsylvania (0.04)
    - Texas > Travis County
      - Austin (0.04)

Genre:
- Research Report > New Finding (1.00)

Industry:
- Education > Educational Setting
  - K-12 Education > Primary School (0.35)
- Health & Medicine > Therapeutic Area
  - Psychiatry/Psychology (0.69)
- Law (1.00)

Technology:
- Information Technology > Artificial Intelligence
  - Machine Learning > Neural Networks
    - Deep Learning (1.00)
  - Natural Language
    - Chatbot (1.00)
    - Large Language Model (1.00)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found