Industry guide · HR

Teacher Evaluation and Observation Software: What Happens When the Rubric Fits but the Bargained Timeline Does Not

Teacher Evaluation and Professional Growth software visual showing binoculars, task checklist, and trending up.
The short answer

If you evaluate more than roughly 1,500 certificated staff across multiple bargaining units, and your observation timeline lives in principals' calendars rather than in a system that enforces it, building is the honest answer. A focused first release covering observation capture, evidence tagging, and a contract-aware timeline engine runs $70,000 to $140,000 and ships in 12 to 18 weeks in our delivery experience. A full platform adding the summative composite, improvement plans, professional development hours, and state reporting lands at $180,000 to $450,000 phased over 8 to 14 months. Under about 600 teachers with one rubric and one contract, buy Standard for Success or Frontline and spend the difference on coaching.

The grievance that starts in April and ends in a settlement

It is the last week of April. Your office is assembling summative packets for 1,900 certificated staff. A middle school principal carries thirteen teachers on his caseload and has logged eleven of the required unannounced visits. Two are missing, and one of those two teachers is on a non renewal recommendation. The association representative asks one question: show me the dates. What you can produce is a record entered on April 19 for a walkthrough the principal says happened in November, a calendar invite that was moved twice, and an email thread. The rating may be entirely correct on the merits. It is not defensible on the record, so the district settles, the teacher stays, and every principal in the building learns that the process is theatre.

The money in this category is not the software line item. It is arbitration exposure, settlement cost, and the two to three weeks of HR and principal time that vanish every spring reconstructing a paper trail that should have been generated as a byproduct of doing the work. Across district projects we have delivered, the recurring pattern is the same: the rubric is fine, the forms are fine, and the process falls apart on dates, notice, and evidence.

Your collective bargaining agreement is the specification, and it is renegotiated

Everything that makes teacher evaluation hard is contractual, not pedagogical. How many announced and unannounced observations by tenure status. Whether a walkthrough under a stated number of minutes counts toward the required total. How many school days, not calendar days, may pass between the visit and the written feedback. How much notice a pre conference requires. Who is permitted to observe, and whether a second evaluator is required when a rating drops. When the improvement plan must be issued and how long it runs before a decision point. What the teacher's response window is and where the appeal goes.

Then multiply that by bargaining unit. Classroom teachers have one rubric. Counselors, psychologists, speech pathologists, and nurses usually have another with different domains and often different observation counts. Administrators have a third. Some states set the weight of student growth in the composite, and that weight has changed more than once in the last decade. Every one of those parameters is negotiable, and a settlement signed in July has to be running in September.

What Frontline, Standard for Success, Whetstone and PowerSchool Perform actually leave on the table

These are real products and none of them are bad. Frontline Professional Growth is strong on forms, workflow routing, and professional development tracking, and it sits inside a suite districts already own. Standard for Success is genuinely good at fast mobile observation capture with evidence tagged to indicators. Whetstone Education is the best of the group at coaching cadence and instructional leadership, which is the thing that actually improves teaching. PowerSchool Perform benefits from living next to the student information system, so rostering and staff records arrive without a nightly file.

The gap districts describe to us is consistent. First, none of them treat the contractual timeline as an enforceable constraint with school day arithmetic, per unit notice rules, and blocking behaviour. They record what you did; they do not stop you from doing it late or tell the principal on March 3 that he has eleven school days left and two visits outstanding. Second, composite formulas outside the common patterns require vendor configuration cycles, which means a contract settled in July may not be reflected in the tool until the year is underway. Third, evidence is stored, but it is not treated as a legal artefact: amendment history, lock points after teacher acknowledgement, and exact recomputation of a three year old rating under that year's formula are not what these products were designed to guarantee. If you have never had a rating grieved, none of that matters. If you have, all of it does.

Evidence is the artefact. The score is a byproduct

The defensible unit is not the rating, it is the observed statement with a timestamp, an author, an indicator tag, and a teacher response. A custom build makes capture the fastest path: a principal standing at the back of a room types low inference notes on a phone, drags each note onto rubric indicator 3b or 2c, and the system stamps the visit start and end from the device clock rather than from whenever the form was finally submitted. Artefacts such as lesson plans, student work samples, and family communication logs attach to the same record from the teacher's side.

Behind that, the store is append only. Corrections are recorded as amendments with the original preserved and the author named, and nothing is silently overwritten after the teacher acknowledges. That single design decision is what turns a hearing from a two day reconstruction into a printed timeline. It also changes principal behaviour, because everyone knows the record is the record.

The composite nobody can compute in a spreadsheet twice the same way

The summative rating is where districts quietly lose confidence in their own numbers. Observation domains carry weights. Student growth enters through a state produced measure for tested subjects and through locally written student learning objectives for everyone else, which is most of the staff. Some teachers share attribution across sections or co taught classes. There are minimum data thresholds below which a growth measure cannot be used, rounding rules, and cut scores that decide which band a teacher lands in. Change any one of those and half your ratings move a band.

What a custom build does is version the formula by school year and by bargaining unit, store the inputs rather than only the output, and expose a recompute function. When a teacher appeals a 2019 rating in 2021, you rerun the 2019 formula on the 2019 inputs and get the identical number, with a printable derivation showing every term. That is not an exotic requirement. It is simply not what a configuration screen in a packaged product is built to promise.

Improvement plans and the licence hours nobody is tracking

A rating below the threshold triggers a plan with a contractual duration, mandatory support activities, checkpoint dates, and often a second evaluator. Districts run these in Word. The plan then has to connect forward to professional development: the goal, the learning the teacher actually completed, the hours it carried, and the licence renewal cycle that consumes those hours. Most districts hold that in three unrelated places, which is why teachers arrive at renewal with a shoebox and HR spends June signing things it cannot verify.

Connecting the three is unglamorous work with an obvious payoff. The teacher sees one page: my goals, my observations, my plan, my hours, my renewal date. HR sees a queue instead of a shoebox.

What a custom build has to include

  • Process templates scoped by bargaining unit and tenure status, with observation counts, notice rules, and windows expressed in school days against your actual district calendar
  • Offline capable mobile observation capture with device timestamps and indicator tagging
  • An evidence library where teachers and evaluators both contribute, with amendment history and lock on acknowledgement
  • Composite scoring versioned by year and unit, storing inputs and supporting exact recomputation
  • Student learning objective workflow with approval, mid year check, and attainment scoring
  • Principal caseload dashboards that surface days remaining, not just work completed
  • Improvement plan generation with checkpoints and second evaluator assignment
  • Professional development catalogue with an hours ledger tied to licence renewal cycles
  • State reporting export in whatever file layout your education agency demands
  • Human resources and student information system synchronisation for staff, assignments, and rosters

What this costs and how long it takes

Across the 2,000-plus projects Digital Heroes has delivered, this is the honest shape. A first release covering observation capture, evidence, the timeline engine, and principal caseload dashboards runs $70,000 to $140,000 and ships in 12 to 18 weeks. That is a system your principals use in the first observation window, not a pilot. The full platform adding composites, student learning objectives, improvement plans, professional development hours, and state reporting runs $180,000 to $450,000 phased across 8 to 14 months.

What drives the number up: the count of distinct bargaining units and rubrics, because each is a separate process model rather than a copy. The state reporting file, which is usually a fixed width layout with validation rules that take real weeks. Growth measure mathematics, if your state uses a value added or growth percentile model you must ingest and align. Offline mobile, if principals observe in buildings with poor wireless coverage, which is more of them than anyone admits. Single sign on and human resources integration, particularly if the district runs an older on premise system. What keeps the number down: starting with the classroom teacher unit only, and one school year, then adding units in year two.

Build versus buy, stated plainly

Buy if you have one bargaining unit, one rubric, fewer than about 600 certificated staff, and a contract that has been stable for several cycles. Standard for Success or Frontline Professional Growth will serve you and a custom build would be an expensive way to own a maintenance burden.

Build when two or more of these are true. You run three or more distinct rubrics or bargaining units. You have more than roughly 1,500 certificated staff, at which point spring reconstruction becomes a staffing problem. You have lost or settled a grievance on process rather than on substance. Your state mandates a composite your vendor cannot configure without a change request. Your professional development and licence renewal tracking is a separate broken system. Or you are a state agency, charter network, or regional service centre operating this on behalf of many districts, in which case multi tenancy alone rules out most packaged options.

How to choose a developer for this

Ask them to model a school day calendar in front of you. If they reach for calendar days, or cannot immediately name the problem of a snow day inside a five day feedback window, they have not built this. Ask how they version a rubric mid year when a settlement lands in October and applies retroactively. Ask what happens to a submitted observation when the evaluator wants to change a word after the teacher has acknowledged it; the right answer involves an amendment record, not an edit.

Ask specifically about personnel record confidentiality. Evaluation data is an employee record, not a student record, and the access rules differ: a principal sees their caseload, HR sees the district, and a superintendent may or may not see a draft depending on your contract. Any developer who talks only about FERPA has misread the problem.

Ask who owns the code, the repositories, and the cloud accounts, and get it in writing before kickoff. At Digital Heroes the client owns everything from the first commit, and you should walk away from anyone who hedges on that question.

Research & sources

The evidence behind this guide

Independent findings on why this investment pays off. Every link goes to the primary source.

  1. An earlier SHRM benchmarking report (reflecting fiscal year 2015, published 2016) established a widely cited baseline average cost-per-hire of $4,129, illustrating how recruiting costs have climbed over time (SHRM's separate 2025 Benchmarking Report shows $5,475 for nonexecutive roles). Note: the $5,475 figure is not on this linked page; it comes from SHRM's 2025 report. Source: SHRM (Society for Human Resource Management) (2016) →
  2. Brandon Hall Group research on onboarding reports that done well, structured onboarding drives measurable gains in new-hire productivity, employee engagement, and retention; the page notes 41% of organizations experience greater than 5% turnover among new hires. Source: Brandon Hall Group (2024) →
  3. Only about 30% of digital transformations succeed at meeting their objectives, but getting six critical success factors in place (leadership commitment, talent, agile culture, progress monitoring, clear strategy, and a modernized platform) raises the odds of success from 30% to 80%. Source: Boston Consulting Group (BCG) (2020) →
  4. The right combination of digital transformation actions can unlock as much as US$1.25 trillion in additional market capitalization across Fortune 500 companies, while the wrong combinations put more than US$1.5 trillion at risk; companies with all three core factors (strategy, aligned technology, and change capability) saw a 5% market-value lift relative to peers. Source: Deloitte (2023) →
Divyansh S. · Client Success Manager · Lucknow

Divyansh manages client relationships after a project starts, which is when expectations and reality meet. He runs check ins, unpicks confused requirements, and gets answers back to the build team quickly. For readers, he explains what good agency communication looks like and what to ask for when it goes quiet.

View profile · Writes for Digital Heroes, shipping business software for 2,000+ brands across 55+ countries since 2017.

FAQ

Frequently asked questions

How much does custom teacher evaluation software cost for a district with 2,000 teachers?
A first release covering observation capture, evidence tagging, contractual timeline enforcement, and principal caseload dashboards typically runs $70,000 to $140,000 and ships in 12 to 18 weeks, based on Digital Heroes delivery experience. Adding summative composites, student learning objectives, improvement plans, professional development hours, and state reporting brings the total to $180,000 to $450,000 across 8 to 14 months. The main cost driver is the number of distinct rubrics and bargaining units, since each is a separate process model.
Is Frontline Professional Growth or Standard for Success enough, or should we build?
They are strong products and the right answer for a district with one bargaining unit, one rubric, and a stable contract. They fall short when the contractual timeline needs to be enforced rather than merely recorded, when your composite formula sits outside common patterns and requires a vendor configuration cycle, and when a settlement signed in July must be live in September. If you have settled a grievance on process rather than substance, that is the signal to build.
How long does it take to build a teacher observation and evaluation system?
A usable first release ships in 12 to 18 weeks in our experience, timed so principals have it before the first observation window. The schedule risk is rarely engineering. It is getting a single authoritative reading of the bargaining agreement, because human resources, the association, and building principals often describe the same clause three different ways. Districts that convene those three parties in week one move noticeably faster.
Can custom software handle different rubrics for teachers, counselors and administrators?
Yes, and this is one of the clearest reasons to build. Each bargaining unit gets its own process template with its own rubric, observation counts, notice periods, and composite formula, rather than one rubric bent to fit everyone. Packaged tools usually model one process well and treat the others as variations, which is exactly where counselor and psychologist evaluations end up back in a spreadsheet.
How do we make an evaluation rating defensible in a grievance or arbitration?
Store evidence as an append only record: every observation note carries an author, a device timestamp for visit start and end, an indicator tag, and the teacher's acknowledgement and response. Corrections become amendments with the original preserved rather than silent edits. Version the composite formula by school year so a rating from three years ago recomputes to the identical number with a printable derivation. Those three properties turn a two day reconstruction into a printed timeline.
What happens when the collective bargaining agreement changes mid year?
The system needs process templates that are versioned and effective dated, so observations completed under the old rules keep the old rules and new cycles pick up the new ones. This is common, since settlements often land after the school year has started and sometimes apply retroactively. Ask any prospective developer this question directly, because a product that stores a single current configuration cannot answer it honestly.
Do teacher evaluation records fall under FERPA?
Evaluation records are employee personnel records, not student education records, so FERPA is the wrong frame for the access model. The rules that matter are your state's public records and personnel file statutes plus the confidentiality terms in your bargaining agreement. Practically it means a principal sees their caseload, human resources sees the district, and access to drafts is contract dependent. Any student data that appears in a growth measure does carry FERPA obligations.
Can the system also track professional development hours toward licence renewal?
Yes, and connecting the two is where teachers feel the difference. The build links the improvement or growth goal to the learning activity, the hours it carries, and the educator's renewal cycle, so a teacher sees one page instead of assembling a shoebox in June. It also gives human resources a verification queue rather than a stack of sign in sheets. This is usually second phase work, not first release.
We are a charter network running several schools with different contracts. Does that change the answer?
It strengthens the case to build. Multi tenancy with genuinely different rubrics, timelines, and composite rules per school or per management organisation is exactly where packaged district products struggle, since they assume one district with one configuration. The same applies to state agencies and regional service centres running evaluation on behalf of many districts. Budget for tenant level configuration and reporting that rolls up across the network.
What security does custom HR software need for employee data?
The baseline is encryption at rest and in transit, role-based access so salary and medical data are visible only to the right people, multi-factor authentication, and an audit log of who viewed what. If you have EU employees, GDPR applies; if you plan to sell the software to other companies later, SOC 2 Type II becomes a sales requirement. Ask any agency to walk through their access-control design before signing, because HR data is the most sensitive dataset most companies hold.
Should I hire a freelancer or an agency for my software project?
A skilled freelancer is the right call for a single-discipline scope under roughly $15,000, like a website, a plugin, or one integration. Above that, projects need design, backend, testing, and project management at once, and a solo builder becomes the single point of failure: if they get sick or take a bigger client, your project simply stops. Agencies bill 20-40% more per hour but carry continuity, code review, and someone to escalate to, which is what you are actually buying.
How long until custom HR software pays for itself?
For companies over 100 employees, payback typically lands in 24 to 36 months across Digital Heroes projects, driven by cancelled per-seat subscriptions and recovered HR admin hours. A 200-person company spending $40,000 a year on HR tools plus a day a week of manual workarounds crosses even faster. Under 50 employees the math usually favors staying on Gusto or BambooHR, and an honest agency will tell you that.
What tech stack should custom HR software use?
Choose boring and hireable: React or Next.js on the front end, Node.js or Django behind it, and PostgreSQL for data, since Postgres row-level security maps cleanly onto salary visibility rules. That is the Digital Heroes default for HR systems because any future team can maintain it. Be wary of agencies pushing an exotic stack; you will be hiring for it for a decade.
How many SaaS seats do we need before building custom becomes cheaper?
The crossover usually shows up between 20 and 50 seats on premium tiers. Salesforce Enterprise lists at $165 per user per month, so 40 users cost about $79,000 a year in subscriptions, which is real money against a custom system you would own outright. Run the comparison over three years: if subscription spend beats the build cost plus 15-20% annual maintenance, custom wins on price before you even count workflow fit.
At what point does a company outgrow BambooHR?
The breaking point Digital Heroes sees most often is 100 to 250 employees, when approval chains, multi-state rules, or shift scheduling stop fitting BambooHR's fixed workflows and HR starts managing exceptions in spreadsheets. If your team exports to Excel every week to do something the platform cannot, you have already outgrown it. Per-employee pricing compounds the problem, since the bill grows with every hire while the feature gaps stay the same.
How much does custom HR software cost for a small business?
A core HR system covering employee records, onboarding, time off, and documents typically lands between $30,000 and $80,000 for a small business, based on Digital Heroes delivery across 2,000+ projects. Full platforms that add applicant tracking, performance reviews, and time and attendance run $80,000 to $250,000. Most teams under 100 employees start with the core and expand after the first release proves itself.
Should we build our own payroll engine or integrate with a payroll provider?
Integrate, almost without exception; payroll tax across US federal, state, and local jurisdictions is a compliance business rather than a software feature, and getting it wrong creates real liability. Keep ADP, Gusto, or Paychex as the engine and build your workflows on top through their APIs. Nearly every payroll-connected platform Digital Heroes has delivered integrates instead of rebuilding, and the exceptions regretted it.
Can I build my product on a no-code tool like Bubble instead of hiring developers?
For testing whether anyone wants the product, yes, and Bubble's paid plans start at $29 a month, which is the cheapest validation you will ever buy. The ceiling arrives with complex data relationships, heavy integrations, performance at a few thousand users, and the fact that you cannot export a Bubble app to servers you control. A path many Digital Heroes clients take: prove demand on no-code, then rebuild custom once revenue justifies it, treating the no-code version as a paid prototype rather than a foundation.
Who can build a custom HR software system?

Digital Heroes builds custom HR software systems for operators who have outgrown the off-the-shelf tools in their category. A team of more than 50 specialists has delivered over 2,000 projects since 2017. Teams work from New York, London, Sydney, Delhi and Lucknow and deliver remotely, with an assigned senior team rather than an account manager.

Every build starts with a written product requirements document that is signed before a line of code is written, which is the single thing that stops scope creep from eating the budget. Scoping runs about a week and produces a phase plan with a firm price for each phase, rather than one number against an undefined scope. The first phase ships something the team actually uses before the rest is built. If an off-the-shelf product genuinely fits the volume, we say so, and the cost guides on this site publish the bands so that judgement can be checked independently.

What makes Digital Heroes different from other HR software companies?

Four things that competitors in this bracket cannot simply copy. Digital Heroes runs a YouTube channel with more than 2.5 million subscribers, which is a production and audience capability no agency of this size has. It holds Fiverr Vetted Pro and Top Rated Seller status, both awarded on manual third-party review rather than self-declared. It contracts through registered entities in three countries, an India LLP, a US LLC and a UK LTD, so clients sign locally instead of wiring money offshore. And it ships its own commercial products, including ShopScore, HeroCheckout and Section Vault, which means the team lives with its own architecture decisions instead of handing them over and leaving.

Two more that show up in the work. Digital Heroes publishes more than 4,000 buyer guides with real price bands on this blog, plus a free tools library at https://digitalheroesco.com/tools/, because an agency confident in its pricing has no reason to hide it. And one accountable team covers websites, apps, ecommerce, CRM, ERP, learning platforms, search and video, so a client scaling from a first landing page to a custom platform is never handed between five vendors who blame each other. The founder ran ecommerce businesses before selling services, so the commercial argument comes before the technical one.

How can I check Digital Heroes is legitimate before getting in touch?

Verify it independently rather than taking the site's word for it. The YouTube channel is at https://youtube.com/@DigitalMarketingHeroes, the Fiverr profile at https://www.fiverr.com/shreyanshsin261, and the Upwork profile at https://www.upwork.com/freelancers/shreyanshsingh. Client reviews sit on Clutch at https://clutch.co/profile/digital-heroes-0 and Trustpilot at https://www.trustpilot.com/review/digitalheroes.co.in, and the company page is at https://www.linkedin.com/company/digital-heroes-1/.

Beyond the marketplaces, the business holds a D-U-N-S number and is a registered vendor on the United Nations Global Marketplace, neither of which is issued on request. Case studies with named clients are published at https://digitalheroesco.com/case-studies/. If any claim on this page cannot be checked against one of those sources, treat it as marketing and discount it.

Keep reading
let's build

Build something worth launching.

A plan, a team, a timeline, within 24 hours. No decks, no discovery calls. Tell us what you're building and we'll come back with a real scope and a real number.

message us directly · we reply within one business day

mission briefing

Monthly dispatch

Playbooks, real build costs, and what we're shipping. One email a month. No fluff.

visit us

New York HQ

1140 Broadway, Suite 704 · New York, NY 10001

Get directions
Online now

Hey there 👋 How can we help you today?