The assistant can guess what you left unsaid. It labels every guess.
Data Mint writes your first codebook in four steps, and you decide how much it may read into what you said.
1
Tell us about your project.
a sentence or two is enough to start
attach what you already have
a codebook from an earlier project
a draft of the one you want
an example document or two
training materials, a coding manual, anything you would hand a new assistant
detail levelTake me literallyRead between the lines
2
Clarifying questions.
it asks four or five, you answer in one box
3
Proposed fields.
one tag and one line per column, not the full text
4
Your codebook.
written out in full: 6 items, 5 hints
Whatever the assistant read between the lines is tagged Hints, so you can audit it field by field.
@karlrohe
Data Mint writes your first codebook in a four-step conversation. You bring whatever you already have. Documents. An old codebook. A draft of the one you want. A finished codebook that is simply not in Data Mint's format. Upload it or paste it in. All of it is useful.
Step 1, from a real run. One box, one sentence, an attach link, and the detail level. The sentence typed here was seventy characters long. The setting under it is the whole of what you decide before the assistant writes anything.
Bring your material. Upload documents, paste text, and describe the data you want.
Answer four or five clarifying questions. The assistant asks for what it is missing.
Review the skeleton. It proposes a tag and a phrase per question, not the full text.
Say go. It writes the codebook out in full.
Step 2, same run. Five questions, then a note recommending Text mode over Scientific or Visual. Questions two and four are the ones that matter here. They ask about cluster randomization and about animal studies, two boundaries the one-sentence description never touched.The answer box below those questions, filled in. Five short replies, typos and all, then a paragraph asking for two things the questions did not offer. Evidence fields before the final decision, and a column that flags anything a human should review. That paragraph is where the codebook actually got its shape.
The skeleton is the cheap place to argue. A tag and a phrase per question is short enough to read in a minute, and short enough to rearrange. Changing the shape of the codebook here costs one sentence. Changing it after the full text is written means editing every field the change touches. So this is the step where I split a double question, ask for a scaffolding question ahead of a hard one, and add what I forgot to say in step 1. Go around as many times as you want.
Step 3, the skeleton. Part A restates the job in one paragraph and asks whether it is right. Part B lists six columns with a line each. The whole screen is shorter than one field of the finished codebook, which is why this is the cheap place to argue.The three doors out of step 3. Edit Directly hands you the text. Reply / Revise sends it back with a note and returns a new skeleton. Looks perfect! moves to step 4. The middle one is the loop, and it has no limit.
The other thing you settle in step 1 is how much the assistant may read into what you said. Two postures. Take me literally writes down what you said and stops there. Read between the lines also fills in what your instructions imply. The second is faster. The first protects validity, because a rule you never settled is not your rule. Underspecification is an instrument, not a defect. Leave an item thin and the ambiguity surfaces as disagreement between readers, where you can see it and decide whether it deserves a policy or stays a residual.
A hint is instruction with its authorship attached. Every sentence the assistant is about to write lands in one of three places. What traces to something you said goes in the Description. What it read between the lines goes in a Hints tag, inside the field it belongs to. What it invented goes nowhere: the prompt forbids writing hypothetical edge cases into a description. So hints are not a staging area. Readers see them at mint time like any other tag. Keeping them in their own tag is what lets you audit them one at a time and cut the ones you disagree with.
Step 4, the output. The banner counts the fields and counts the hints separately. Below it the codebook is plain text, frontmatter first, then one section per field. Five hints across six fields is the number to look at. It is how much of this codebook the assistant wrote rather than transcribed.
definitions
posture — how literally the codebook assistant takes your instructions. The screen calls it Detail level. You pick Take me literally or Read between the lines in step 1, before it writes anything.
hint — a line the assistant wrote to answer a question your own description raised. Tagged Hints so you can find it. Examples stays your material only.
Nothing here is settled until you have read it. The tags tell you who wrote each line. They do not tell you which lines you agree with. A hint you never opened is still in the codebook when the readers run. What you have at the end of the conversation is a draft with its authorship marked. The deciding starts when readers disagree.