01
Social Discourse Index
Salience and stance of key topics across regions and time.
- Coverage
- US, FR, ES, GB, DE, IT, +7
- Frequency
- Weekly
- History
- 12 months
- Unit
- region × topic × week
Clamor Data transforms social video into structured datasets across topic, entity, sentiment, stance, geography and time.
Dataset catalogue
| Dataset | Coverage | Frequency | Labels | History |
|---|---|---|---|---|
| Social Discourse Index | US, FR, ES, GB, DE, IT, +7 | Weekly | topic · entity · stance · sentiment · region | 12 months |
| Political Narrative | US, GB | Daily | entity · issue · stance · narrative · sentiment | 18 months |
| Consumer Intent | US, CA, AU, FR, DE, ES | Weekly | category · intent · sentiment · brand | 12 months |
| Economic Sentiment | US, EU | Weekly | topic · sentiment · intensity · region | 24 months |
Topic salience over time
United States · weekly
Our pipeline
Speech is separated from catalogue audio, transcribed, resolved against a maintained taxonomy and aggregated by region and week. Every stage is measured, and the sample size travels with the record.
Video sources
Transcription
Entity & topic
Stance & sentiment
Region & time
Structured dataset
Sample record
Each record is time-stamped, geo-attributed and labeled across multiple dimensions — the same shape whether it arrives over the API, in a Parquet drop or in a historical backfill.
01
Salience and stance of key topics across regions and time.
02
Political entities, issues and narratives at scale.
03
Category demand, brand mentions and purchase intent.
04
Public sentiment around economic and financial topics.
| Market | Languages | Videos processed | History | Cadence |
|---|---|---|---|---|
| United States | en, es | 21.4M | 24 months | Daily |
| United Kingdom | en | 6.1M | 18 months | Daily |
| Spain | es, ca | 4.8M | 12 months | Weekly |
| France | fr | 5.2M | 12 months | Weekly |
| Germany | de | 4.4M | 12 months | Weekly |
| Italy | it | 3.0M | 12 months | Weekly |
Query by region, topic, entity and date range. Cursor pagination, ISO-8601 timestamps, versioned endpoints and a stable record contract.
Parquet and CSV drops to S3, GCS or Azure on your cadence. Full historical backfill lands with the first delivery.
Bespoke taxonomies, entity lists and regional cuts built against a specific research question, delivered on the same schema.
Quality assurance
Sampling is uniform over the identifier space, so a share is a share of a measurable population rather than of whatever a search returned.
Live records of known state are interleaved with every collection run, which separates 'absent' from 'blocked' instead of silently under-counting.
Every aggregate ships with n and a 95% interval. Rows thinner than the stated floor are published as null rather than smoothed.
A stratified slice of each release is hand-checked against the transcript, and the agreement rate is published with the release notes.
Tell us the market and the question. We return a real slice — schema, sample sizes and confidence intervals included — not a brochure.