Skip to content

Consulting and recruiting

Market research firm data inventory: records by system

By SourceX Editorial · Updated

Short answer

A market research data inventory lists every system that holds the firm's records and, for each one, the record types, the consent basis respondents gave, whether the client owns the material and the retention rule. Start with the survey platform, recruitment database, recording and transcription tools, analysis files and shared drives, and record metadata only.

Key takeaways

  • One row per system per record type keeps consent and ownership answers precise.
  • Consent basis and client ownership are the two columns that decide what the firm can do with a record.
  • Collaboration tools scatter research files across mailboxes, OneDrive and SharePoint, so list each location.
  • Fixed values such as firm-owned, client-owned or unknown make rows easy to filter and compare.
  • The inventory describes records; it never needs copies of them.

What a market research data inventory records#

A market research data inventory is a metadata list: each row names a system, a type of record it holds and the facts that decide how that record can be used. It copies no data, so it can circulate among the COO, compliance and IT without exposing respondent or client material.

Research firms need more columns than most businesses because their records carry promises in two directions: to respondents, through consent and privacy notices, and to clients, through master agreements and statements of work. The template keeps both visible on every row, so nobody has to remember which studies came with which strings.

The template: systems, record types and key columns#

Start with these rows and adjust the systems to your own stack. The values shown illustrate how to fill each column; they are not recommendations for your firm.

The template: systems, record types and key columns
SystemRecord typesConsent basisClient ownershipRetention
Survey platformQuestionnaires, raw responses, quotas, quality flagsStudy-specific consent at survey startResponses often client-controlled; questionnaires varyRaw files to report delivery plus a set period
Recruitment databaseScreener answers, contact details, participation historyDatabase membership termsFirm-ownedWhile membership is active
Recording platform and facility portalVideo and audio of groups and interviewsSession consent formOften a client deliverableShort; delete after analysis or handover
Transcription vendor portalIdentified transcriptsSession consent formFollows the recordingDelete once a de-identified version exists
Analysis and tabulation filesCleaned datasets, weights, table specsInherited from the studyMixed; check the SOWPer contract
Coding toolCodeframes and coded verbatimsInherited from the studyCodeframes usually firm-ownedPer contract
Norms databaseStudy-level scores and metadataNot respondent-levelFirm-owned under aggregated-use clausesBy written recency rule
CRM and proposal libraryProposals, pricing, client contactsNot respondent dataFirm-ownedPer firm policy
Project management toolTimelines, tasks, issue logsNot respondent dataFirm-ownedPer firm policy
Email and Teams or SlackClient threads and shared filesVaries by attachmentMixedPer mail and chat policy

Use a short list of fixed values for the two judgment columns so rows can be sorted and compared. Free text drifts quickly; a fixed list forces each question to be answered or marked unknown.

Unknown is an acceptable first answer. A row marked unknown shows exactly where the next contract or notice review is needed, which is more useful than a confident guess that later turns out wrong.

  • Consent basis: study-only consent; consent includes further research use; membership terms allow reuse; not respondent data; unknown.
  • Client ownership: firm-owned; client-owned deliverable; client confidential but usable in aggregate; mixed by project; unknown.
  • Retention: name the trigger event and the source of the period, such as report delivery plus the contract period.
  • Notes: record exceptions, such as a client that prohibits third-party AI tools on its studies.

Supporting columns worth adding#

The five core columns decide what a record can be used for. A few supporting columns make the inventory useful for planning a migration, answering a client audit or scoping any later licensing review, and each one can be filled from a conversation with the system owner.

Supporting columns worth adding
ColumnWhat to recordWhy it helps
Date rangeMonth and year of the earliest and latest recordsShows depth of history and where consent versions change
Rough volumeStudies, sessions or files, as an order of magnitudeSizes any export or deletion job
Export routeNative export, API, vendor request or noneReveals records that are hard to retrieve
Linked recordsWhat each record connects to, such as study ID or client IDLinked records are easier to trace and to remove
LanguageMain language of open-ends and transcriptsMatters for coding tools and outside use
System ownerNamed person who answers for the systemKeeps the row current after staff changes

Systems research firms often forget#

The systems most often missed are the ones where research files land by accident. Collaboration tools are the biggest gap. Microsoft states that Teams chat messages are stored in a hidden folder in the mailbox of each user in the chat, and channel messages in a similar hidden folder in a group mailbox, reached through eDiscovery rather than directly. Files shared in Teams usually end up in personal OneDrive or team SharePoint storage, so a data file dropped into a chat may not sit in the project folder at all.

  • Personal OneDrive and desktop folders where analysts keep working copies.
  • Old survey platform accounts left open after a move to a new platform.
  • Transcription and translation vendor portals that still hold files.
  • Facility streaming and video portals from in-person groups.
  • Incentive and payment platforms holding respondent contact and payment details.
  • Freelancers' and subcontractors' own storage.
  • Email attachments carrying data files and banner tables.

How to build the inventory without moving files#

Build the inventory from interviews and admin screens, not exports. Each step uses information that system owners can give from what they already see, so no respondent or client file has to leave its system.

  • List systems from the IT asset register, single sign-on and expense reports.
  • Interview each system owner about record types, the date range and rough volume.
  • Pull consent versions from survey and screener templates, with the dates each version was in use.
  • Pull ownership answers from master agreements and a sample of statements of work.
  • Copy retention rules from the retention schedule, and note any tool that deletes automatically.
  • Review the draft with operations, compliance and one senior researcher before calling it final.

Illustrative: a B2B research firm maps its records#

Illustrative: a fictional research firm runs B2B studies for software and industrial clients, combining online surveys, a recruited database of professionals and remote in-depth interviews. The COO's first draft lists the obvious systems: the survey platform, the recruitment database and the shared drive.

Interviews with system owners add an old survey platform account still holding several years of studies, a transcription portal nobody had closed and a Teams channel where data files were exchanged with a fieldwork partner. Each becomes its own row.

The consent column shows that recent interviews used a form with a separate choice about reuse in methods research, while older ones did not. The ownership column shows that most master agreements allow aggregated use of results. The firm takes the exports it is allowed to keep, closes the stale accounts and schedules a contract review for rows marked unknown.

How SourceX uses an inventory like this#

SourceX starts every engagement with metadata, which is exactly what this inventory holds. In the Supply step of the SourceX five-step transaction, the inventory shows which record families exist, how far back they go and which rows are blocked by consent or client ownership.

The SourceX Enterprise Data Value Framework then rates the remaining rows qualitatively, on drivers such as uniqueness, domain expertise, human-generated signal, recency, rights and privacy burden. Nothing is shared during the initial assessment, and the firm approves every later step.

Frequently asked questions

Who should own the inventory?

Usually the COO or operations lead, with compliance owning the consent and ownership columns and IT supplying the system list. One named owner keeps it current; shared ownership tends to mean nobody updates it after the first review.

Should client-owned records be listed?

Yes. The inventory maps what the firm holds, and client-owned material is part of that. Listing it with the right ownership value prevents accidental reuse and makes client audits and return-or-destroy requests faster to answer.

Should we list data held by fieldwork partners and panel suppliers?

List what they hold on your behalf, such as files they host for your studies or recordings in their portals, with the partner named as the location. Their own panel databases belong in a separate note, because those records are governed by the partner's terms with its members, not by yours.

How often should the inventory be updated?

Update it when a system is added or retired, when consent or contract templates change and at each retention review. A dated version history shows how the firm's records have changed, which helps in client audits and any due diligence.

Is this the same as a GDPR record of processing?

They overlap but serve different purposes. A record of processing activities under the GDPR focuses on purposes, legal bases and transfers of personal data. This inventory also covers non-personal records and client ownership. Many firms keep both and cross-reference them.

Sources

  • Microsoft states that Teams chat data is stored in a hidden folder in the mailbox of each user included in the chat, and channel messages use a similar hidden folder in a group mailbox. Source

Related resources

See if your company qualifies

A short company assessment. No data uploads are needed.

See if you qualify