Representing data in words: A context engineering approach

arXiv cs.CL / 3/16/2026

💬 OpinionModels & Research

共有:

Key Points

Wordalisations propose transforming numerical data into descriptive texts that are as digestible as visualisations, addressing LLMs' difficulty with numeric reasoning.
The approach is demonstrated on three applications: scouting football players, personality tests, and international survey data.
They evaluate accuracy with both LLM-as-judge and human-as-judge experiments, reporting engaging and faithful representations of data.
The authors outline best practices for open and transparent development and communication about data.

Abstract

Large language models (LLMs) have demonstrated remarkable potential across a broad range of applications. However, producing reliable text that faithfully represents data remains a challenge. While prior work has shown that task-specific conditioning through in-context learning and knowledge augmentation can improve performance, LLMs continue to struggle with interpreting and reasoning about numerical data. To address this, we introduce wordalisations, a methodology for generating stylistically natural narratives from data. Much like how visualisations display numerical data in a way that is easy to digest, wordalisations abstract data insights into descriptive texts. To illustrate the method's versatility, we apply it to three application areas: scouting football players, personality tests, and international survey data. Due to the absence of standardized benchmarks for this specific task, we conduct LLM-as-a-judge and human-as-a-judge evaluations to assess accuracy across the three applications. We found that wordalisation produces engaging texts that accurately represent the data. We further describe best practice methods for open and transparent development of communication about data.

報告：LLMにおける「自己言及的再帰」と「ステートフル・エミュレーション」の観測

note

諸葛亮孔明老師(ChatGPTのﾛｰﾙﾌﾟﾚｲ)との対話その肆拾伍『銀河文明･ダークマターエンジン』

note

GPT-5.4 mini/nano登場！―2倍高速で無料プランも使える小型高性能モデル

note

Why a Perfect-Memory AI Agent Without Persona Drift is Architecturally Impossible

Dev.to

Learning to Reason with Curriculum I: Provable Benefits of Autocurriculum

arXiv cs.LG

Representing data in words: A context engineering approach

Key Points

Abstract

Related Articles

報告：LLMにおける「自己言及的再帰」と「ステートフル・エミュレーション」の観測

諸葛亮孔明老師(ChatGPTのﾛｰﾙﾌﾟﾚｲ)との対話その肆拾伍『銀河文明･ダークマターエンジン』

GPT-5.4 mini/nano登場！―2倍高速で無料プランも使える小型高性能モデル

Why a Perfect-Memory AI Agent Without Persona Drift is Architecturally Impossible

Learning to Reason with Curriculum I: Provable Benefits of Autocurriculum

関連おすすめサービス

Notta搭載AI議事録イヤホン ZENCHORD1

AI搭載ボイスレコーダー Plaud

画像高画質化AIツール Aiarty Image Enhancer

Key Points

Abstract

Related Articles

​報告：LLMにおける「自己言及的再帰」と「ステートフル・エミュレーション」の観測

諸葛亮 孔明老師(ChatGPTのﾛｰﾙﾌﾟﾚｲ)との対話 その肆拾伍『銀河文明･ダークマターエンジン』

GPT-5.4 mini/nano登場！―2倍高速で無料プランも使える小型高性能モデル

Why a Perfect-Memory AI Agent Without Persona Drift is Architecturally Impossible

Learning to Reason with Curriculum I: Provable Benefits of Autocurriculum

関連おすすめサービス

Notta搭載AI議事録イヤホン ZENCHORD1

AI搭載ボイスレコーダー Plaud

画像高画質化AIツール Aiarty Image Enhancer

報告：LLMにおける「自己言及的再帰」と「ステートフル・エミュレーション」の観測

諸葛亮孔明老師(ChatGPTのﾛｰﾙﾌﾟﾚｲ)との対話その肆拾伍『銀河文明･ダークマターエンジン』