• Skip to main content
  • Skip to primary sidebar
  • Skip to footer
  • Core Principles of Responsible AI
    • Accountability
    • Fairness
    • Privacy
    • Safety and Security
    • Sustainability
    • Transparency
  • Special Topics
    • AI in Industry
    • Ethical Implications
    • Human-Centered Design
    • Regulatory Landscape
    • Technical Methods
  • Living Dictionary
  • State of AI Ethics
  • AI Ethics Brief
  • 🇫🇷
Montreal AI Ethics Institute

Montreal AI Ethics Institute

Democratizing AI ethics literacy

From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting

December 14, 2023

🔬 Research Summary by Griffin Adams, a final year NLP PhD student at Columbia University under Noémie Elhadad and Kathleen McKeown, who will be starting as the Head of Clinical NLP for Stability AI in 2024.

[Original paper by Griffin Adams, Alexander R. Fabbri, Faisal Ladhak, Eric Lehman, and Noémie Elhadad]


Overview: Selecting the “right” amount of information to include in a summary is difficult: a good summary should be detailed and entity-centric without being overly dense and hard to follow. To better understand this tradeoff, we solicit increasingly dense GPT-4 summaries with what we refer to as a “Chain of Density” (CoD) prompt. Specifically, GPT-4 generates an initial entity sparse summary before iteratively incorporating missing salient entities without increasing the length.


Introduction

Automatic summarization has come a long way in the past few years, largely due to a paradigm shift away from supervised fine-tuning on labeled datasets to zero-shot prompting with Large Language Models (LLMs), such as GPT-4 (OpenAI, 2023). Careful prompting can enable fine-grained control over summary characteristics, such as length, topics, and style, without additional training. An overlooked aspect is the information density of a summary. Theoretically, as a compression of another text, a summary should be denser–containing a higher concentration of information–than the source document. Given the high latency of LLM decoding, covering more information in fewer words is a worthy goal, especially for real-time applications. Yet, how dense is an open question. A summary is uninformative if it contains insufficient detail. However, if it contains too much information, it can become difficult to follow without increasing the overall length. Conveying more information subject to a fixed token budget requires a combination of abstraction, compression, and fusion. There is a limit to how much space can be made for additional information before becoming illegible or even factually incorrect.

Key Insights

In this paper, we seek to identify the optimal balance between detail and readability by using GPT-4 to generate increasingly entity-dense (e.g., detailed) summaries and have humans provide preference assessments. 

The Prompt

The Chain of Density (CoD) prompt is achieved with a single prompt to GPT-4, which is tasked with writing five summaries of a provided article. At each step, 1-3 additional details (entities) are added to the previous summary without increasing the length. Existing content is re-written to make room for new entities (e.g., compression, fusion).

The Data

We randomly sample 100 articles from a CNN/DailyMail news article collection.

Human Feedback

We conduct a human evaluation to assess the impact of densification on human assessments of overall quality. Specifically, the first four authors of the paper were presented with randomly shuffled CoD summaries, along with the articles, for the same 100 articles (5 steps * 100 = 500 total summaries). Based on the same definition of a “good summary,” each annotator indicated their top preferred summary.  Our results indicated that humans prefer summaries that are almost as dense as human-written summaries and more dense than summaries generated from a simple GPT-4 prompt: “Write a VERY short summary of the Article. Do not exceed 70 words.”

Between the lines

We study the impact of summary densification on human preferences for overall quality. A degree of densification is preferred, yet it is very difficult to maintain readability and coherence when summaries contain too many entities per token. We open-source annotated test sets and a larger unannotated training set for further research into the topic of fixed-length, variable-density summarization.  Future work should identify the optimal information level to include for each unique article.  Given the rise of open-source LLMs (LLama, Mistral), this expensive Chain of Density prompt could be distilled into a single model through fine-tuning.

Want quick summaries of the latest research & reporting in AI ethics delivered to your inbox? Subscribe to the AI Ethics Brief. We publish bi-weekly.

Primary Sidebar

🔍 SEARCH

Spotlight

AI Policy Corner: Singapore’s National AI Strategy 2.0

AI Governance in a Competitive World: Balancing Innovation, Regulation and Ethics | Point Zero Forum 2025

AI Policy Corner: Frontier AI Safety Commitments, AI Seoul Summit 2024

AI Policy Corner: The Colorado State Deepfakes Act

Special Edition: Honouring the Legacy of Abhishek Gupta (1992–2024)

related posts

  • Research summary: Changing My Mind About AI, Universal Basic Income, and the Value of Data

    Research summary: Changing My Mind About AI, Universal Basic Income, and the Value of Data

  • Algorithmic Auditing and Social Justice: Lessons from the History of Audit Studies

    Algorithmic Auditing and Social Justice: Lessons from the History of Audit Studies

  • Research summary: Artificial Intelligence: The Ambiguous Labor Market Impact of Automating Predictio...

    Research summary: Artificial Intelligence: The Ambiguous Labor Market Impact of Automating Predictio...

  • Generative AI-Driven Storytelling: A New Era for Marketing

    Generative AI-Driven Storytelling: A New Era for Marketing

  • The social dilemma in artificial intelligence development and why we have to solve it

    The social dilemma in artificial intelligence development and why we have to solve it

  • Research summary: Algorithmic Colonization of Africa

    Research summary: Algorithmic Colonization of Africa

  • Research summary: Integrating ethical values and economic value to steer progress in AI

    Research summary: Integrating ethical values and economic value to steer progress in AI

  • AI Ethics Maturity Model

    AI Ethics Maturity Model

  • Embedding Ethical Principles into AI Predictive Tools for Migration Management in Humanitarian Actio...

    Embedding Ethical Principles into AI Predictive Tools for Migration Management in Humanitarian Actio...

  • Combatting Anti-Blackness in the AI Community

    Combatting Anti-Blackness in the AI Community

Partners

  •  
    U.S. Artificial Intelligence Safety Institute Consortium (AISIC) at NIST

  • Partnership on AI

  • The LF AI & Data Foundation

  • The AI Alliance

Footer

Categories


• Blog
• Research Summaries
• Columns
• Core Principles of Responsible AI
• Special Topics

Signature Content


• The State Of AI Ethics

• The Living Dictionary

• The AI Ethics Brief

Learn More


• About

• Open Access Policy

• Contributions Policy

• Editorial Stance on AI Tools

• Press

• Donate

• Contact

The AI Ethics Brief (bi-weekly newsletter)

About Us


Founded in 2018, the Montreal AI Ethics Institute (MAIEI) is an international non-profit organization equipping citizens concerned about artificial intelligence and its impact on society to take action.


Archive

  • © MONTREAL AI ETHICS INSTITUTE. All rights reserved 2024.
  • This work is licensed under a Creative Commons Attribution 4.0 International License.
  • Learn more about our open access policy here.
  • Creative Commons License

    Save hours of work and stay on top of Responsible AI research and reporting with our bi-weekly email newsletter.