Start from the standard, not the feature list. The tool has to take your achievement standard, marking schedule and success criteria as its input, then draft comments you can edit. Judge it on whether its grade judgements would survive moderation in your subject, and on where student data is stored.
Details about other products were checked on . Other providers change their features and pricing, so check their site before deciding.
What should an AI marking tool actually do for NCEA work?
The job is narrower than most product demos suggest. A marking tool should take three inputs: the student’s work, the standard you are assessing against, and the marking schedule you would have used anyway. It should hand back a draft comment and a provisional grade judgement, in language close to your own, that you can accept, edit or throw out. If a tool cannot take your marking schedule as an input, it is a rubric generator wearing a marking label. That single distinction decides most of the value you will get.
NCEA sharpens the test, because an internal assessment decision does not stop at your desk. NZQA describes internal moderation as the process that ensures assessment is valid and that grade judgements are verified, and external moderation as assurance that assessment decisions in relation to assessment standards are consistent nationally. Its assessor support material goes further, letting assessors practise judgements on full samples of student work and compare them with grades awarded by moderation panels. So the honest question about any AI tool is not whether the comments read well. It is whether the judgement it drafts near a grade boundary is one you would defend to a moderator.
That gives you three things to check, in order. Does it read the script accurately. Does it apply your criteria rather than a generic set. Does it point to the evidence in the work that justifies what it said. A tool that does the first two well and the third badly will cost you more time than it saves, because you end up re-reading every script to work out why it landed where it did.
Is a broad AI suite or a focused marking tool the better fit?
Two shapes of product dominate. The first is the broad teaching suite, and MagicSchool is the best known example. It advertises more than 80 teacher tools and more than 50 student tools, including a Lesson Plan Generator, Rubric Generator, Worksheet Generator, Report Card Comments, IEP Generator and Presentation Generator, plus integrations with Google, Microsoft, Canvas, Schoology, Clever and ClassLink. There is a free plan aimed at individual teachers, a Plus plan listed at USD $8.33 a month billed annually or USD $12.99 month to month, and a custom-priced Enterprise plan that adds SSO, SIS and LMS integrations and custom data privacy agreements. Its student side runs on Student Rooms, with teacher visibility, built-in guardrails and content filtering.
That breadth is a genuine strength and worth saying plainly. If your Sunday goes on building resources, drafting a unit, writing report comments or turning a reading into a differentiated worksheet, a wide suite covers all of it in one login, and a free tier means you can test that claim this week at no cost. Nothing about a specialist marking tool beats it on that ground.
The second shape is the dedicated marking assistant, which does one job and carries the machinery that job needs: your rubric and success criteria as the source of truth, comment banks you reuse, and the same criteria applied across a whole class so that script thirty is judged the way script one was. The trade-off is real in both directions. Breadth wins when planning is the bottleneck. Depth wins when the bottleneck is a stack of internals due Friday and the risk you are managing is drift in your own judgement across a long marking session.
- Pick a broad suite if planning, resources and communication are eating the most hours.
- Pick a dedicated marking tool if consistent, criteria-aligned marking is the bigger load.
- Run both if your week is genuinely split, and keep the free tier of the suite as your planning layer.
How do you test a tool against your own marking schedule?
No feature list can answer this. If you are weighing up an AI marking tool for NCEA, run a small trial on work you have already marked. Take one internal assessment from last year that has been through moderation. Pull six scripts that sit across a grade boundary, remove names and any identifying detail, and give the tool your marking schedule verbatim rather than asking it to invent one. Then compare its draft judgement and its reasoning against the grades you and your faculty already agreed on. Six scripts is enough to see a pattern and small enough to do in a free period.
Judge the trial on agreement near the boundary, not on average agreement. Any competent model will get a clear top-band script and a clear bottom-band script right. The money is in the middle. Keep the trial separate from any student-facing use such as chat tutors or study guides, because those raise a different set of questions about supervision and consent that deserve their own decision.
Write down what you find before you buy anything. If a tool disagrees with your faculty consistently in one direction, that is useful information about how you would have to prompt it, or a reason to walk away.
- Does it work from the marking schedule you supply, or does it substitute its own rubric?
- Does it quote or point to evidence in the script for each judgement?
- Does it stay consistent when the same script is submitted twice?
- Does it hold your subject vocabulary, or does it drift into generic praise?
- Can you save success criteria and comment banks and reuse them across a class?
What should a New Zealand school check before student work leaves the building?
Privacy is where these decisions are usually won or lost, and it is a school decision rather than a teacher one. Your privacy officer or board will want to know what leaves the school, where it lands, who can see it and how long it is kept, and they will want it in the contract rather than on a marketing page. Start that conversation before a trial begins, not after a class set has already been uploaded.
On the named example, here is what the vendor’s own documentation says as at 29 July 2026. MagicSchool is a US company and its privacy policy states that data is accessed and processed in the United States. It states that it does not use personal information to train artificial intelligence or machine learning models, that third-party AI providers are contractually prohibited from using processed data for model development, and that data retained by those services is deleted within thirty days. It states that it complies with the EU-US Data Privacy Framework and its UK extension, and that it uses standard contractual clauses for transfers where no adequacy decision applies. Custom data privacy agreements are listed as an Enterprise feature rather than a default. None of that is a problem in itself, but US processing is a fact your school needs to weigh against its own obligations, and vendor terms change, so read the current version before you sign.
Ask the same questions of every vendor, including the ones with a local accent. A tool built closer to home is not automatically compliant, and a US-hosted tool is not automatically ruled out.
- Where is student work stored and processed, and can you choose the region?
- Is student work ever used to train or improve the models?
- How long is data retained by the vendor and by any AI provider behind it?
- Which independent audits or certifications can the vendor actually produce?
- Does the agreement meet your school's Privacy Act and procurement requirements?
Where does a dedicated marking assistant like Jeddle fit?
Full disclosure before this section: this is our own tool, so read it as a description of the category rather than a recommendation. Jeddle is built in Australia and used across Australian and New Zealand schools, and it sits squarely in the second shape described above. You supply the rubric, success criteria and comment banks, JeddAI drafts marks and comments against them, and you review and edit everything before a student sees it. The design intent is consistency across a class and less repetitive typing, not a shortcut around your judgement, which is why the routine comments get drafted and the borderline scripts still land on your desk.
It does not try to be a planning suite, and it will not write your worksheets or run a student chat tutor. If your load is heaviest at the point of expert essay marking against a standard, run the same six-script trial described earlier and check the output against the grades your faculty already agreed. If your load is heaviest at the planning end, a broad suite is the better first purchase, and there is no reason the two cannot coexist.
| What you are comparing | Broad suite (example: MagicSchool) | Dedicated marking assistant |
|---|---|---|
| Breadth | 80+ teacher tools and 50+ student tools | One workflow: marking and feedback |
| Where criteria come from | Built-in generators such as the Rubric Generator, outputs editable | The marking schedule and success criteria you already use |
| Entry cost | Free plan for individual teachers; Plus listed at USD $8.33/month billed annually | Varies by vendor; check current plans |
| Student-facing use | Student Rooms with teacher visibility and content filtering | Usually teacher-side only |
| Data location | US company; data accessed and processed in the United States | Ask the vendor for stated residency in writing |
| Model training | States it does not use personal information to train AI models | Ask for the same commitment in writing |
Frequently asked questions
Is MagicSchool free?
There is a free plan for individual teachers that includes the teacher and student tools, quizzes and class writing feedback. A Plus plan is listed at USD $8.33 a month billed annually, or USD $12.99 month to month, and Enterprise pricing is custom. Prices are in US dollars, so check the pricing page for current figures.
Can a general AI tool mark against NCEA achievement standards?
No mainstream tool is purpose-built for NCEA, so the test is whether it will work from the marking schedule you supply rather than a rubric it generates. Internal moderation still verifies your grade judgements, and NZQA external moderation checks them nationally.
Does AI marking replace a teacher's judgement?
No. These tools draft output that you review and edit before a student sees it, and the mark you report is yours. That matters most near grade boundaries, where a best-fit judgement is exactly the part a moderator will look at.
Which type of tool is better for lesson planning?
A broad suite. MagicSchool's lesson plan, worksheet, quiz and presentation generators cover planning far more widely than any marking-focused tool, and the free tier makes it cheap to test against your own units.
Can I use a planning tool and a marking tool together?
Yes, and many teachers do. The two categories rarely overlap much in practice, so pairing a free or low-cost planning suite with a dedicated marking tool is a common setup. Just run each through your school's privacy process.
Get started with Jeddle
Jeddle gives teachers and students instant, syllabus-aligned feedback powered by JeddAI.
Looking for study material? Browse Jeddle's Australian-English subject resources, or explore more articles on EdTech & Tools.



