TELPAS Writing Rubric: The Twelve-Point Scale, Collections, and Annotated Examples

Telo AI helps school districts improve speaking outcomes for English Learners and support bilingual education programs through conversational AI and practical tools for educators.

Middle-school Emergent Bilingual student writing at a desk while a teacher leans in to review the page

Estimated reading time: 10 minutes

The TELPAS writing rubric is what a Texas teacher applies when judging an Emergent Bilingual student’s written English, and there is more than one. Grades 4 through 12 have a twelve-point constructed-response rubric scoring three traits, while holistic writing ratings across grades 2 through 12 are judged against the proficiency level descriptors from the ELPS. Districts that confuse the two produce inconsistent ratings and cannot explain them afterwards. This guide covers both, what a writing collection has to contain, how to read TEA’s annotated student samples, and why writing is the domain where a district’s effort translates most directly into a rating.

Table of contents

Executive Summary

Texas assesses TELPAS writing through student writing collected in the classroom rather than through a single timed essay, which makes the rubric a year-long instrument rather than a test-day one. For grades 4 through 12, TEA publishes a twelve-point constructed-response rubric that scores three traits, vocabulary, usage, and completeness, each on a one-to-four scale. Alongside it, holistic writing ratings are assigned against the proficiency level descriptors drawn from the English Language Proficiency Standards. TEA also publishes annotated examples of student writing, which are the single most useful free resource a district has for aligning judgment across campuses. Writing rewards preparation more reliably than any other domain, because the evidence is produced in the normal course of instruction and can be improved deliberately.

Key Takeaways

  • Grades 4-12 use a twelve-point constructed-response rubric: vocabulary, usage, completeness, each scored 1 to 4.
  • Holistic writing ratings are judged against the ELPS proficiency level descriptors.
  • Writing is collected across the year, not produced in a single sitting.
  • TEA publishes annotated student writing examples, the best free calibration resource available.
  • Writing is the most improvable domain, because the evidence is a normal classroom product.
  • Collection quality drives rating quality. A thin collection cannot be rated well.

The Twelve-Point Writing Rubric, Grades 4-12

Quick answer: TEA’s constructed-response writing rubric for grades 4 through 12 scores three traits, vocabulary, usage, and completeness, on a one-to-four scale each, for a maximum of twelve points.

The three-trait structure is worth understanding rather than memorising, because it tells a teacher what to teach. The rubric is not scoring whether the student had a good idea. It is scoring whether the English carried it.

TraitWhat it asksWhat a low score usually means
VocabularyIs the word choice sufficient and precise for the task?Reliance on a narrow set of high-frequency words
UsageIs grammar controlled enough not to obstruct meaning?Errors that force the reader to reconstruct intent
CompletenessIs the response developed enough to show what the student can do?A response too short to demonstrate range

The published rubric anchors each level in the effect on comprehensibility rather than in error counts. At the lowest level it describes widespread spelling errors that significantly interfere with comprehensibility and significant grammar usage errors that interfere. At the highest it describes infrequent spelling errors that do not interfere, and grammar usage that is generally correct and comparable to that of grade-level native English-speaking peers.

That phrasing is the key to applying it consistently. A rater is not counting mistakes; they are asking how hard the reader has to work. Two students with the same number of errors can sit at different levels if one set of errors obstructs meaning and the other does not.

Why Completeness Is the Trait Districts Underestimate

Vocabulary and usage are what teachers expect to be scored. Completeness is where ratings quietly go wrong.

A student who writes three careful sentences with no errors has not demonstrated much. The rubric cannot credit range that was never attempted, so a cautious, accurate, very short response caps out low. This produces a counterintuitive instructional message that is nonetheless correct: for TELPAS writing purposes, a student attempting a longer piece with some errors often demonstrates more than a student producing a short flawless one.

Teaching to that reality is legitimate rather than gaming. Extended writing is what the standards ask for, so encouraging Emergent Bilinguals to write at length, and treating errors as evidence of reach rather than of failure, improves both the instruction and the rating.

What a Writing Collection Has to Contain

TELPAS writing in the collected grades is assembled from work students produce in class over the year, which is what separates it from an on-demand essay test. The practical requirements a district should build its year around are these.

  1. Multiple pieces per student, gathered across the year rather than in one push, so the collection reflects typical performance.
  2. A range of writing types, so a student is not judged solely on the one genre they happen to handle best.
  3. Authentic classroom work, produced under normal conditions rather than staged for the collection.
  4. Content-area writing, not only language-arts writing, which matters because the ELPS make language development every teacher’s responsibility.
  5. Documentation that survives a challenge, since the rating informs reclassification decisions a parent may question.

Confirm the exact composition requirements for the current administration in the TEA TELPAS materials, because collection specifications are revised between cycles. The strategic point holds regardless of the year’s detail: the collection is built in September, not in February.

Use TEA’s Annotated Student Writing

The most underused free resource in Texas EB assessment is TEA’s annotated examples of student writing. These are real student pieces with the state’s reasoning attached, explaining why a given sample sits where it sits.

A district that runs a one-hour session comparing its own teachers’ ratings of the annotated samples against the state’s annotations will learn more about its rating consistency than a year of completion reports. Where campuses diverge, they diverge predictably: some raters over-weight spelling, others credit effort, and the annotations settle both arguments with an authority no internal moderator has. Rater alignment more broadly is covered in TELPAS calibration and holistic rating.

Funding Writing Instruction and Rater Alignment

Funding sourceEligible usesHow to access
Bilingual Education AllotmentDirect programme costs; at least 10 percent on EB professional developmentState formula, Texas Education Code section 48.105
Title III, Part ASupplemental EB instruction, technology, and trainingFormula grant through TEA
Title I, Part AAcademic support in high-poverty campusesFormula grant through TEA
TIPS and regional cooperativesPre-approved purchasing of qualifying toolsCooperative contract, often no separate RFP

Writing moderation sessions are EB professional development and therefore fundable from the allotment’s 10 percent training floor. See Title III funding in Texas for the federal side.

District Benchmark: The Collection Arithmetic

Writing collections are the part of TELPAS a district can actually control, and the volume is manageable if it is planned.

Take a district with 2,000 Emergent Bilinguals, of whom roughly 1,600 sit in the collected grades. If each student contributes several pieces across the year, the district is handling several thousand pieces of student writing that have to be gathered, stored, associated with the right student, and available to a rater in February. Spread across 36 school weeks that is a light, continuous task. Compressed into three weeks it is chaos, and the ratings show it.

The comparison worth making internally is with the speaking domain. Writing evidence accumulates as a by-product of teaching, so a district that simply organises collection gets a defensible sample almost for free. Speaking evidence does not accumulate that way, which is why the two domains diverge in every TELPAS report in the state.

Why Writing Outruns Speaking in Every District Report

Read any Texas district’s TELPAS data by domain and writing sits above speaking. The cause is not that writing is easier to learn. It is the Speaking Time Gap: the shortfall between the responsive spoken practice an Emergent Bilingual needs and what one teacher can supply across a full class.

The math, run across the two productive domains. Writing and speaking are both production skills, and both are assessed from work a student generates. The difference is parallelism. In a 45-minute block, all 25 students can write at once, so a single teacher can generate 25 pieces of evidence simultaneously and collect 1,125 student-minutes of production. In that same block, spoken production happens one student at a time, so the ceiling is a couple of minutes each no matter how the lesson is run. Writing scales with the class; speaking divides by it.

That single structural difference explains the domain gap in the data, and it also explains why writing responds to a district’s effort while speaking resists it. A district that decides to improve writing can do so inside the existing schedule. A district that decides to improve speaking has to find capacity the schedule does not contain.

Metrics Worth Tracking

  • Pieces collected per Emergent Bilingual by December, as an early warning on collection health.
  • Share of collected writing from content areas rather than language arts alone.
  • Rating agreement against TEA’s annotated samples, by campus.
  • Average response length for Emergent Bilinguals, given the completeness trait.
  • The writing-to-speaking rating gap, tracked year over year as a capacity indicator.

Common Mistakes District Leaders Make

  1. Collecting only language-arts writing. The ELPS make every teacher responsible for language development.
  2. Compressing collection into the weeks before the window. It produces an unrepresentative sample.
  3. Rewarding short and safe writing. The completeness trait cannot credit range that was not attempted.
  4. Never using the annotated examples. They resolve rater disagreements no internal discussion can.
  5. Reading the writing rating in isolation. Its distance from the speaking rating is the more useful number.

Immediate (this month): Run a one-hour moderation session using TEA’s annotated student writing and record how far each campus sits from the state’s ratings.

Medium-term (this year): Put a collection calendar in place from September, with content-area writing built in, and fund the moderation sessions from the Bilingual Education Allotment training requirement.

Long-term (strategy): Track the writing-to-speaking gap as a standing indicator of instructional capacity, and resource the speaking side, which the schedule cannot supply on its own.

Questions District Leaders Should Ask

  • How many pieces of writing does each Emergent Bilingual have collected by December?
  • What proportion of that writing comes from outside language arts?
  • How closely do our campuses agree with TEA’s annotated ratings?
  • Are our Emergent Bilinguals writing at length, or safely and briefly?
  • How wide is the gap between our writing and speaking ratings, and is it narrowing?

Frequently Asked Questions

What is the TELPAS writing rubric?

For grades 4 through 12, TEA publishes a twelve-point constructed-response rubric that scores three traits, vocabulary, usage, and completeness, on a one-to-four scale each. Holistic writing ratings are separately judged against the proficiency level descriptors drawn from the English Language Proficiency Standards.

What three traits does the TELPAS writing rubric score?

Vocabulary, usage, and completeness. Each is scored from 1 to 4, giving a maximum of 12 points. The descriptors are anchored in how far errors interfere with comprehensibility rather than in how many errors appear.

What is a TELPAS writing collection?

It is a set of writing a student produces in class across the year, gathered so that raters judge typical performance rather than a single timed piece. Because it is assembled continuously, the quality of the collection is decided in September, not in February.

Where can I find TELPAS writing examples?

TEA publishes annotated examples of student writing showing real student work with the state’s reasoning attached. They are the most useful free resource for aligning judgment between raters and campuses.

What kind of writing prompts should we use?

A range, and not only from language arts. Content-area writing belongs in the collection because the ELPS make language development every teacher’s responsibility, and a collection drawn from one subject shows only what a student can do in that subject.

Does length matter in TELPAS writing?

Yes, through the completeness trait. A short flawless response cannot demonstrate range, so it caps out lower than a longer response that attempts more and contains some errors. Encouraging Emergent Bilinguals to write at length improves both the instruction and the rating.

Is there an equivalent rubric for speaking?

Yes, speaking is rated against its own descriptors. It is covered separately in our guide to the TELPAS speaking rubric, because the two domains are rated differently and, in most districts, produce very different results.

Conclusion

The TELPAS writing rubric rewards districts that plan, which makes writing unusual among the four domains. The criteria are published, the annotated exemplars are free, and the evidence is produced by ordinary teaching. A district that organises collection from September, moderates its raters against TEA’s own annotations, and teaches students to write at length rather than safely will see the writing rating move. The more informative number is what that leaves behind: the distance between writing and speaking, which measures not how well the district teaches but how much individual attention its schedule can physically deliver.

Sources: TEA, TELPAS Twelve-Point Writing Rubric for Grades 4 through 12; TEA, TELPAS Annotated Examples of Student Writing; Texas Education Agency, TELPAS.

The Domain That Answers, and the Domain That Does Not

A Texas district that takes writing seriously usually sees it. Organise the collection, moderate the raters against the state’s exemplars, push students to write at length, and the writing ratings climb. It is one of the few places in EB programme management where effort and outcome line up cleanly.

Which makes the contrast painful. The same district applies the same seriousness to speaking and the ratings barely move, because writing can be produced by 25 students at once and speech cannot. The gap is arithmetic, not commitment.

So districts are looking for speaking capacity that does not come out of a teacher’s minute-by-minute attention. Telo AI is one example, offering adaptive conversation in English plus Spanish and French for bilingual programmes.

See how districts close the gap between their writing and speaking data: https://mytelo.ai/how-telo-works/

A rubric describes what good work looks like. It cannot create the conditions in which a student produces it. Writing gets those conditions for free from a normal school day; speaking has to be given them deliberately.

Found this guide useful?

Preferred sources show up more often in your Google results, including AI Overviews.

Add Telo AI as a preferred source in Google

See how Telo works in the classroom

Learn how Telo helps English Learners practice speaking at their own level while giving teachers real-time insights.