<?xml version="1.0" encoding="utf-8"?>
<feed xmlns="http://www.w3.org/2005/Atom">
  <title>Amila Sampath / notes on design &amp; data</title>
  <subtitle>Notes by Amila Sampath on design, data and government websites in Sri Lanka, from building Overlook, an independent audit of 575 government websites on top of the Lanka Data Foundation scorecard.</subtitle>
  <link href="https://www.govlk.site/blog/"/>
  <link rel="self" href="https://www.govlk.site/blog/feed.xml"/>
  <id>https://www.govlk.site/blog/</id>
  <updated>2026-10-07T00:00:00Z</updated>
  <author><name>Amila Sampath</name><uri>https://www.govlk.site/blog/</uri></author>
  <entry>
    <title>What 575 Sri Lankan government websites tell us, and what AI got right</title>
    <link href="https://www.govlk.site/blog/what-575-government-websites-tell-us/"/>
    <id>https://www.govlk.site/blog/what-575-government-websites-tell-us/</id>
    <published>2026-10-05T00:00:00Z</published>
    <updated>2026-10-07T00:00:00Z</updated>
    <summary>Amila Sampath built Overlook on top of Lanka Data Foundation's volunteer scorecard to audit 575 Sri Lankan government websites with a browser robot and AI. Phone speed is the biggest problem, and after three scans the automated scorecard matched the volunteers' grade on 75.4% of sites.</summary>
    <category term="GovTech"/>
    <category term="Sri Lanka"/>
    <category term="UX research"/>
    <category term="Accessibility"/>
    <category term="Data visualisation"/>
    <category term="AI"/>
    <content type="html">  &lt;figure class=&quot;wide&quot;&gt;
    &lt;img src=&quot;https://www.govlk.site/blog/what-575-government-websites-tell-us/img/overlook-wall.jpg&quot; alt=&quot;The Overlook board: a grid of Sri Lankan government home pages on the left, sorted by LDF score, with government design system reference pages from the UK, US, Canada, Australia, NSW and Estonia on the right. Header stats read 556 of 575 sites shown, average LDF score 69, median Lighthouse performance 32 and accessibility 85.&quot; width=&quot;1600&quot; height=&quot;1000&quot;&gt;
    &lt;figcaption&gt;Overlook, the tool this post is about. Every government home page on one zoomable wall, with other governments' design systems in a column on the right. &lt;a href=&quot;https://www.govlk.site&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Open it at www.govlk.site&lt;/a&gt;.&lt;/figcaption&gt;
  &lt;/figure&gt;

  &lt;div class=&quot;wrap&quot;&gt;
    &lt;h2 id=&quot;start&quot;&gt;Where this started&lt;/h2&gt;
    &lt;p&gt;In September, &lt;a href=&quot;https://opendata.lk/&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lanka Data Foundation&lt;/a&gt; (LDF) ran its Government Website Scorecard Digital Audit. On 5 September 2026, 24 volunteers sat down together and checked &lt;span class=&quot;num&quot;&gt;577&lt;/span&gt; government websites against a simple, public rubric: does the site load, does it have a valid SSL certificate, can you read it in Sinhala, Tamil and English, does it use the official gov.lk domain, and does its contact email match that domain. They published the results on 25 September.&lt;/p&gt;
    &lt;p&gt;I loved it. It is the kind of civic research that rarely happens because it is slow and boring to do by hand. It also left me with a question I could not shake: how much of this could one person do alone with today's AI tools, and could they go further than the rubric?&lt;/p&gt;
    &lt;p&gt;So I tried. The first version took me about eight hours. Everything since has been about checking whether the numbers are right.&lt;/p&gt;

    &lt;div class=&quot;facts&quot; role=&quot;list&quot;&gt;
      &lt;div role=&quot;listitem&quot;&gt;&lt;strong&gt;575&lt;/strong&gt;&lt;span&gt;sites in the LDF scorecard list I worked from&lt;/span&gt;&lt;/div&gt;
      &lt;div role=&quot;listitem&quot;&gt;&lt;strong&gt;24&lt;/strong&gt;&lt;span&gt;LDF volunteers, one day, by hand&lt;/span&gt;&lt;/div&gt;
      &lt;div role=&quot;listitem&quot;&gt;&lt;strong&gt;1&lt;/strong&gt;&lt;span&gt;person, about 8 hours for the first build&lt;/span&gt;&lt;/div&gt;
      &lt;div role=&quot;listitem&quot;&gt;&lt;strong&gt;34,769&lt;/strong&gt;&lt;span&gt;screenshots and crops on the live board&lt;/span&gt;&lt;/div&gt;
    &lt;/div&gt;

    &lt;h2 id=&quot;tool&quot;&gt;What I built: Overlook&lt;/h2&gt;
    &lt;p&gt;Overlook takes LDF's list of sites and opens every one in a real browser, first as a desktop at &lt;span class=&quot;num&quot;&gt;1280 × 800&lt;/span&gt;, then as a phone at &lt;span class=&quot;num&quot;&gt;390 × 844&lt;/span&gt;. On each site it:&lt;/p&gt;
    &lt;ul&gt;
      &lt;li&gt;takes the first screen as a citizen lands on it, clicks through any &quot;choose your language&quot; page to the English home page, then visits one inner page;&lt;/li&gt;
      &lt;li&gt;finds the common parts of the page (header, menu, search, hero, footer, logo, page layout) and crops each one out of the live page;&lt;/li&gt;
      &lt;li&gt;runs Lighthouse for speed on a phone, plus axe-core and HTML_CodeSniffer for accessibility, and its own checks for Sinhala and Tamil language tags, tap targets and skip links;&lt;/li&gt;
      &lt;li&gt;tries the components the way a person would: types &quot;contact&quot; into search and reads what comes back, opens the phone menu, tabs through the main menu with a keyboard, and watches the hero slider for motion and a pause button.&lt;/li&gt;
    &lt;/ul&gt;
    &lt;p&gt;Then it lays everything out side by side on one big zoomable wall. You can sort by LDF score, group by ministry or site type, and click any tile for that site's full sheet. A &quot;See as&quot; mode shows any page the way someone with colour blindness, low vision or cataracts would see it.&lt;/p&gt;
    &lt;p&gt;Next to our sites sits a column of government design systems: GOV.UK, the US Web Design System, Canada, Australia's AgDS, NSW and Estonia's TEDI, rebuilt from their published code. When you look at &lt;span class=&quot;num&quot;&gt;258&lt;/span&gt; search boxes from our estate, you can see right away what a well-specified one looks like.&lt;/p&gt;

    &lt;h3&gt;Ask the data&lt;/h3&gt;
    &lt;p&gt;The wall answers &quot;what does it look like&quot;. People kept asking &quot;so what does it mean&quot;. So I added a chat that sits on top of everything the tool measured. You ask in plain words, like &quot;Which site types are slowest on a phone?&quot;, and get a short answer with a live chart. Click a bar and the wall opens on those exact sites.&lt;/p&gt;
    &lt;p&gt;Because this is public data about public institutions, I spent most of my time on guard rails rather than features. Every number in an answer is checked against what the database actually returned before you see it, and any figure the AI cannot back up is removed. Charts are re-run with the data tables emptied, so a number typed in by the model, rather than counted, gets caught. It costs well under one US cent per question.&lt;/p&gt;
  &lt;/div&gt;

  &lt;figure class=&quot;wrap&quot;&gt;
    &lt;video controls preload=&quot;none&quot; poster=&quot;https://www.govlk.site/blog/what-575-government-websites-tell-us/img/video-cover.jpg&quot; width=&quot;1280&quot; height=&quot;720&quot; aria-label=&quot;A 60-second walkthrough of asking Overlook a question and opening a shared dashboard&quot;&gt;
      &lt;source src=&quot;https://www.govlk.site/blog/what-575-government-websites-tell-us/img/walkthrough.mp4&quot; type=&quot;video/mp4&quot;&gt;
    &lt;/video&gt;
    &lt;figcaption&gt;A 60-second walkthrough: ask a question, get a chart, open the sites behind it, turn it into a shareable dashboard.&lt;/figcaption&gt;
  &lt;/figure&gt;

  &lt;div class=&quot;wrap&quot;&gt;
    &lt;h2 id=&quot;findings&quot;&gt;What it found&lt;/h2&gt;
    &lt;p&gt;The LDF rubric is about basics, and the estate does well on some of them. &lt;span class=&quot;num&quot;&gt;95.1%&lt;/span&gt; of online sites have a valid SSL certificate. Languages are where the points are lost: only &lt;span class=&quot;num&quot;&gt;51.2%&lt;/span&gt; offer a language switcher. Once I looked past the rubric, at what a citizen actually experiences on a phone, the picture changed a lot.&lt;/p&gt;

    &lt;h3&gt;Speed on a phone is the biggest problem&lt;/h3&gt;
    &lt;p&gt;Of the &lt;span class=&quot;num&quot;&gt;494&lt;/span&gt; sites Lighthouse could test, the median mobile performance score is &lt;span class=&quot;num&quot;&gt;32&lt;/span&gt; out of 100. Only &lt;span class=&quot;num&quot;&gt;5&lt;/span&gt; score 90 or more. On a simulated mid-range phone, a typical home page takes about &lt;span class=&quot;num&quot;&gt;15&lt;/span&gt; seconds to show its main content, and only &lt;span class=&quot;num&quot;&gt;7&lt;/span&gt; sites do it within the 2.5-second &quot;good&quot; mark.&lt;/p&gt;
  &lt;/div&gt;

  &lt;figure class=&quot;wrap&quot;&gt;
    &lt;div class=&quot;chart&quot;&gt;
      &lt;h4&gt;Most government sites score under 50 for speed on a phone&lt;/h4&gt;
      &lt;p class=&quot;sub&quot;&gt;Lighthouse 13 mobile performance score, 494 sites, grouped in tens. Captured 30 September 2026.&lt;/p&gt;
      &lt;div class=&quot;legend&quot;&gt;&lt;span&gt;&lt;i class=&quot;leg-soft&quot;&gt;&lt;/i&gt;0–49 (poor)&lt;/span&gt;&lt;span&gt;&lt;i class=&quot;leg-navy&quot;&gt;&lt;/i&gt;50–89&lt;/span&gt;&lt;span&gt;&lt;i class=&quot;leg-saf&quot;&gt;&lt;/i&gt;90–100 (good)&lt;/span&gt;&lt;/div&gt;
      &lt;div class=&quot;scroll&quot;&gt;&lt;/div&gt;
    &lt;/div&gt;
  &lt;/figure&gt;

  &lt;div class=&quot;wrap&quot;&gt;
    &lt;p&gt;A good scorecard grade does not mean a fast site: &lt;span class=&quot;num&quot;&gt;84%&lt;/span&gt; of grade-A sites still score under 50 on phone performance. The scorecard and the experience measure different things, and both matter.&lt;/p&gt;

    &lt;h3&gt;The same component, built many different ways&lt;/h3&gt;
    &lt;ul&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;258&lt;/span&gt; home-page search boxes come in &lt;span class=&quot;num&quot;&gt;181&lt;/span&gt; different sizes. Search shows up in &lt;span class=&quot;num&quot;&gt;7&lt;/span&gt; different forms, and &lt;span class=&quot;num&quot;&gt;224&lt;/span&gt; of &lt;span class=&quot;num&quot;&gt;519&lt;/span&gt; home pages have no search at all.&lt;/li&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;53&lt;/span&gt; of &lt;span class=&quot;num&quot;&gt;490&lt;/span&gt; phone menu buttons did not open when tapped.&lt;/li&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;124&lt;/span&gt; home-page carousels move on their own. Only &lt;span class=&quot;num&quot;&gt;1&lt;/span&gt; has a pause button.&lt;/li&gt;
      &lt;li&gt;Only &lt;span class=&quot;num&quot;&gt;16.6%&lt;/span&gt; of sites have a skip link, something every design system gives you ready-made.&lt;/li&gt;
    &lt;/ul&gt;

    &lt;h3&gt;Languages and accessibility&lt;/h3&gt;
    &lt;ul&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;12%&lt;/span&gt; of working sites open on a &quot;choose your language&quot; gate before you see any content.&lt;/li&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;93%&lt;/span&gt; of pages with Sinhala or Tamil text do not tag that text as Sinhala or Tamil, so screen readers read it with the wrong voice, or not at all.&lt;/li&gt;
      &lt;li&gt;The median Lighthouse accessibility score is &lt;span class=&quot;num&quot;&gt;85&lt;/span&gt;.&lt;/li&gt;
    &lt;/ul&gt;

    &lt;h3&gt;Security and housekeeping&lt;/h3&gt;
    &lt;ul&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;260&lt;/span&gt; of &lt;span class=&quot;num&quot;&gt;545&lt;/span&gt; reachable sites send none of the six standard security headers.&lt;/li&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;230&lt;/span&gt; sites load trackers before the visitor has agreed to cookies.&lt;/li&gt;
      &lt;li&gt;Home pages weigh &lt;span class=&quot;num&quot;&gt;7.7 MB&lt;/span&gt; on average.&lt;/li&gt;
    &lt;/ul&gt;

    &lt;p class=&quot;pull&quot;&gt;Only &lt;span&gt;7&lt;/span&gt; sites clear every basic: online, valid SSL, gov.lk, a language switcher, accessibility 90+ and phone speed 50+.&lt;/p&gt;

    &lt;h2 id=&quot;human-vs-ai&quot;&gt;The real test: would the machine agree with the volunteers?&lt;/h2&gt;
    &lt;p&gt;Pictures and Lighthouse scores are one thing. The harder question was whether automation could reproduce LDF's human scorecard on the same rubric. If it could, the scorecard could be re-run every month instead of once a year. If it could not, I wanted to know exactly where and why.&lt;/p&gt;
    &lt;p&gt;I wrote a crawler that opens each site's core pages (Home, About, Services, Contact, RTI, News), switches into Sinhala and Tamil, counts which script each page is really written in, checks the SSL certificate on every address a citizen might type, reads the contact emails and applies the domain rules from LDF's report. Then I compared its score with the volunteers' score for every site.&lt;/p&gt;

    &lt;h3&gt;The first attempt was not good&lt;/h3&gt;
    &lt;p&gt;On the first run, the scan matched the volunteers' grade on only &lt;span class=&quot;num&quot;&gt;35%&lt;/span&gt; of sites. My first instinct was to blame the volunteers. When I looked closer, most of the gap was my own scan's fault. It counted any Sinhala word in the page code as a language switcher, credited full translation when a switcher merely offered a language without opening it, judged every email on six pages when volunteers judge the main one, and checked the certificate on the bare domain when many certificates only cover the www address.&lt;/p&gt;
    &lt;p&gt;So I rebuilt it to behave more like a volunteer: only count language controls a person can see, actually open each language version, and treat a machine translation widget as &quot;dynamic&quot; the way the rubric does. Then I did it again, adding language menus that open on click, image buttons, splash pages and bot walls.&lt;/p&gt;
  &lt;/div&gt;

  &lt;figure class=&quot;wrap&quot;&gt;
    &lt;div class=&quot;pair&quot;&gt;
      &lt;div class=&quot;chart&quot;&gt;
        &lt;h4&gt;Three scans, closer each time&lt;/h4&gt;
        &lt;p class=&quot;sub&quot;&gt;Share of sites where the scan's grade matches the volunteers' grade.&lt;/p&gt;
        &lt;div class=&quot;scroll&quot;&gt;&lt;/div&gt;
      &lt;/div&gt;
      &lt;div class=&quot;chart&quot;&gt;
        &lt;h4&gt;Agreement by rubric item, first scan to final scan&lt;/h4&gt;
        &lt;p class=&quot;sub&quot;&gt;Share of sites where the scan and the volunteers gave the same points.&lt;/p&gt;
        &lt;div class=&quot;legend&quot;&gt;&lt;span&gt;&lt;i class=&quot;leg-soft&quot;&gt;&lt;/i&gt;First scan&lt;/span&gt;&lt;span&gt;&lt;i class=&quot;leg-saf&quot;&gt;&lt;/i&gt;Final scan&lt;/span&gt;&lt;/div&gt;
        &lt;div class=&quot;scroll&quot;&gt;&lt;/div&gt;
      &lt;/div&gt;
    &lt;/div&gt;
    &lt;figcaption&gt;Scans 1 and 2 cover all 575 listed sites; scan 3 covers the 513 sites both the volunteers and the scan could load, which is the fair comparison. r is the correlation between the two scores.&lt;/figcaption&gt;
  &lt;/figure&gt;

  &lt;div class=&quot;wrap&quot;&gt;
    &lt;p&gt;On the final scan, across the &lt;span class=&quot;num&quot;&gt;513&lt;/span&gt; sites both sides could load, the grades match exactly on &lt;span class=&quot;num&quot;&gt;75.4%&lt;/span&gt; of sites and are within one grade on &lt;span class=&quot;num&quot;&gt;93.2%&lt;/span&gt;. The average score is &lt;span class=&quot;num&quot;&gt;72.8&lt;/span&gt; for the scan and &lt;span class=&quot;num&quot;&gt;72.9&lt;/span&gt; for the volunteers, the correlation is &lt;span class=&quot;num&quot;&gt;0.85&lt;/span&gt;, and the typical gap is &lt;span class=&quot;num&quot;&gt;5.4&lt;/span&gt; points. Only &lt;span class=&quot;num&quot;&gt;6&lt;/span&gt; sites are three or more grades apart.&lt;/p&gt;
  &lt;/div&gt;

  &lt;figure class=&quot;wrap&quot;&gt;
    &lt;div class=&quot;chart square&quot;&gt;
      &lt;h4&gt;Volunteer score against automated score, per site&lt;/h4&gt;
      &lt;p class=&quot;sub&quot;&gt;513 sites both could load. Dots on the dashed line got the same score from both. Hover a dot for the site.&lt;/p&gt;
      &lt;div class=&quot;scroll&quot;&gt;&lt;/div&gt;
    &lt;/div&gt;
  &lt;/figure&gt;

  &lt;div class=&quot;wrap&quot;&gt;
    &lt;h3&gt;Who was right when they disagreed?&lt;/h3&gt;
    &lt;p&gt;Agreement numbers hide the interesting part, so I went through every disagreement. For the ones the evidence could not settle, I took screenshots with images on, clicked the language controls by hand and traced the network requests. There were &lt;span class=&quot;num&quot;&gt;381&lt;/span&gt; category-level disagreements across &lt;span class=&quot;num&quot;&gt;274&lt;/span&gt; sites.&lt;/p&gt;
  &lt;/div&gt;

  &lt;figure class=&quot;wrap&quot;&gt;
    &lt;div class=&quot;chart&quot;&gt;
      &lt;h4&gt;Most disagreements are judgement calls, not mistakes&lt;/h4&gt;
      &lt;p class=&quot;sub&quot;&gt;Verdicts on 381 category-level disagreements, after manual review on 5 October 2026.&lt;/p&gt;
      &lt;div class=&quot;scroll&quot;&gt;&lt;/div&gt;
    &lt;/div&gt;
  &lt;/figure&gt;

  &lt;div class=&quot;wrap&quot;&gt;
    &lt;p&gt;More than half are judgement calls: the rubric does not say exactly which pages count as &quot;core pages&quot; or how much Sinhala makes a page &quot;translated&quot;, and volunteers scored identical platforms differently. The volunteers were wrong more often than the scan (&lt;span class=&quot;num&quot;&gt;72&lt;/span&gt; against &lt;span class=&quot;num&quot;&gt;27&lt;/span&gt;). A few examples:&lt;/p&gt;
    &lt;ul&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;12&lt;/span&gt; sites from one ministry's section were published &lt;span class=&quot;num&quot;&gt;20&lt;/span&gt; points below the sum of their own criteria, because the translation points were dropped. irrigation.gov.lk shows &lt;span class=&quot;num&quot;&gt;79&lt;/span&gt; where its own criteria add up to &lt;span class=&quot;num&quot;&gt;99&lt;/span&gt;.&lt;/li&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;24&lt;/span&gt; sites are marked &quot;email matches the domain&quot; while their contact address is a Gmail or other-domain address.&lt;/li&gt;
      &lt;li&gt;&lt;span class=&quot;num&quot;&gt;21&lt;/span&gt; disagreements are simply sites that changed between 5 September and our scan.&lt;/li&gt;
    &lt;/ul&gt;
    &lt;p&gt;This is not a criticism of the volunteers. Twenty-four people scoring 577 sites in one day will make slips, which is exactly why a second, automated pass is useful. The point is that the two work best together.&lt;/p&gt;

    &lt;p class=&quot;pull&quot;&gt;Automation is reliable on the deterministic checks. It struggles where a person has to &lt;span&gt;read&lt;/span&gt;: is this page really in Tamil?&lt;/p&gt;

    &lt;p&gt;SSL, domain and switcher checks now agree on &lt;span class=&quot;num&quot;&gt;95&lt;/span&gt; to &lt;span class=&quot;num&quot;&gt;97.5%&lt;/span&gt; of sites. The weakest item is still the count of core pages in all three languages, at &lt;span class=&quot;num&quot;&gt;59.3%&lt;/span&gt;, and that is where the rubric itself needs a tighter definition before any method, human or machine, can be consistent.&lt;/p&gt;

    &lt;h2 id=&quot;dashboards&quot;&gt;Three dashboards, three audiences&lt;/h2&gt;
    &lt;p&gt;Once the data was trustworthy, the next question from people was &quot;what should I do with it?&quot;. The chat can turn any answer into a shareable dashboard with one click. No login, and everyone sees the same charts. I made three ready-made ones, each with &lt;span class=&quot;num&quot;&gt;13&lt;/span&gt; or &lt;span class=&quot;num&quot;&gt;14&lt;/span&gt; charts, and every chart opens the real sites behind it.&lt;/p&gt;
  &lt;/div&gt;

  &lt;div class=&quot;wide dashes&quot;&gt;
    &lt;a class=&quot;dash&quot; href=&quot;https://www.govlk.site/#dash=897sks&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;
      &lt;img src=&quot;https://www.govlk.site/blog/what-575-government-websites-tell-us/img/dash-designers.jpg&quot; alt=&quot;The designers dashboard on govlk.site&quot; loading=&quot;lazy&quot; width=&quot;1440&quot; height=&quot;900&quot;&gt;
      &lt;div&gt;&lt;b&gt;For designers&lt;/b&gt;&lt;span&gt;What citizens meet on the home page: carousels, search, skip links, layouts and components compared with design systems.&lt;/span&gt;&lt;u&gt;govlk.site/#dash=897sks&lt;/u&gt;&lt;/div&gt;
    &lt;/a&gt;
    &lt;a class=&quot;dash&quot; href=&quot;https://www.govlk.site/#dash=d7zz62&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;
      &lt;img src=&quot;https://www.govlk.site/blog/what-575-government-websites-tell-us/img/dash-developers.jpg&quot; alt=&quot;The developers dashboard on govlk.site&quot; loading=&quot;lazy&quot; width=&quot;1440&quot; height=&quot;900&quot;&gt;
      &lt;div&gt;&lt;b&gt;For developers&lt;/b&gt;&lt;span&gt;Speed, page weight, security headers, platforms and CMS versions across the estate.&lt;/span&gt;&lt;u&gt;govlk.site/#dash=d7zz62&lt;/u&gt;&lt;/div&gt;
    &lt;/a&gt;
    &lt;a class=&quot;dash&quot; href=&quot;https://www.govlk.site/#dash=rsajkz&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;
      &lt;img src=&quot;https://www.govlk.site/blog/what-575-government-websites-tell-us/img/dash-managers.jpg&quot; alt=&quot;The managers dashboard on govlk.site, showing key findings, scorecard grades and rubric compliance&quot; loading=&quot;lazy&quot; width=&quot;1440&quot; height=&quot;900&quot;&gt;
      &lt;div&gt;&lt;b&gt;For managers&lt;/b&gt;&lt;span&gt;The whole estate at a glance: grades by site type, compliance, and a risk matrix of what to fix first.&lt;/span&gt;&lt;u&gt;govlk.site/#dash=rsajkz&lt;/u&gt;&lt;/div&gt;
    &lt;/a&gt;
  &lt;/div&gt;

  &lt;div class=&quot;wrap&quot;&gt;
    &lt;h2 id=&quot;govtech&quot;&gt;Why this matters for GovTech&lt;/h2&gt;
    &lt;p&gt;Sri Lanka's digital government work is moving to &lt;a href=&quot;https://govtech.lk/&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;GovTech Sri Lanka&lt;/a&gt;, the state company the Cabinet approved in 2025 to &lt;a href=&quot;https://www.ft.lk/front-page/GovTech-to-take-over-ICTA-role/44-776948&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;take over ICTA's role&lt;/a&gt; and lead national digital transformation. Projects like GovPay and a single citizen entry point depend on government websites people can actually use: on a phone, in their own language, and safely.&lt;/p&gt;
    &lt;p&gt;I think this experiment is useful to that work in four ways:&lt;/p&gt;
    &lt;ul&gt;
      &lt;li&gt;&lt;strong&gt;A baseline.&lt;/strong&gt; The estate now has a measured starting point beyond the rubric: speed, accessibility, security headers, components and languages, dated and repeatable.&lt;/li&gt;
      &lt;li&gt;&lt;strong&gt;Fixes that help many sites at once.&lt;/strong&gt; Many problems come from shared templates and the same few platforms. Fix the template and dozens of sites improve together.&lt;/li&gt;
      &lt;li&gt;&lt;strong&gt;A case for a shared design system.&lt;/strong&gt; 181 sizes of search box is what happens without one. Every system on the right-hand column of the wall shows a tested, accessible alternative, and the wall shows which patterns our sites already share.&lt;/li&gt;
      &lt;li&gt;&lt;strong&gt;A cheaper scorecard.&lt;/strong&gt; The rules-based checks agree with volunteers on 95% or more of sites. They can run every month, leaving people free for the parts that need human judgement.&lt;/li&gt;
    &lt;/ul&gt;

    &lt;div class=&quot;note&quot;&gt;
      &lt;h3&gt;What this is, and what it is not&lt;/h3&gt;
      &lt;p&gt;Overlook is an independent experiment. It is not an official audit and not a Government of Sri Lanka or LDF product. The LDF scores, grades and ranks are LDF's published numbers from 5 September 2026. Screenshots and components were captured on 29 September 2026, Lighthouse ran on 30 September (mobile, simulated throttling, performance moves about 5 to 10 points between runs), and the human-vs-AI comparison is from 5 October 2026. Sites change all the time, so any single number may be out of date for a given site. Chats on the board are recorded so I can improve the answers.&lt;/p&gt;
    &lt;/div&gt;

    &lt;h2 id=&quot;lessons&quot;&gt;What I learned&lt;/h2&gt;
    &lt;ul&gt;
      &lt;li&gt;&lt;strong&gt;One person can now do quantitative research across hundreds of sites.&lt;/strong&gt; The building is fast. The checking is the real work, and it should take most of your time.&lt;/li&gt;
      &lt;li&gt;&lt;strong&gt;When the machine disagrees with people, look at your own method first.&lt;/strong&gt; My first scan was wrong far more often than the volunteers were.&lt;/li&gt;
      &lt;li&gt;&lt;strong&gt;Humans and automation are better together.&lt;/strong&gt; The scan caught volunteer slips; the volunteers set the standard the scan had to learn.&lt;/li&gt;
      &lt;li&gt;&lt;strong&gt;Showing beats telling.&lt;/strong&gt; A wall of 575 home pages makes the case for a shared design system faster than any table.&lt;/li&gt;
    &lt;/ul&gt;
    &lt;p&gt;Thank you to Lanka Data Foundation and its volunteers for doing the hard work first. None of this would exist without their list and their scorecard. If you work on government websites, at GovTech, in a ministry or as a vendor, I would love to hear what would make this more useful to you.&lt;/p&gt;

    &lt;h2 id=&quot;links&quot;&gt;Links&lt;/h2&gt;
    &lt;div class=&quot;links&quot;&gt;
      &lt;a href=&quot;https://www.govlk.site&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;&lt;b&gt;Overlook&lt;/b&gt;&lt;span&gt;The board, the chat and the dashboards&lt;/span&gt;&lt;i&gt;www.govlk.site&lt;/i&gt;&lt;/a&gt;
      &lt;a href=&quot;https://opendata.lk/&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;&lt;b&gt;Lanka Data Foundation&lt;/b&gt;&lt;span&gt;The volunteer scorecard this builds on&lt;/span&gt;&lt;i&gt;opendata.lk&lt;/i&gt;&lt;/a&gt;
      &lt;a href=&quot;https://govtech.lk/&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;&lt;b&gt;GovTech Sri Lanka&lt;/b&gt;&lt;span&gt;The government's digital transformation agency&lt;/span&gt;&lt;i&gt;govtech.lk&lt;/i&gt;&lt;/a&gt;
    &lt;/div&gt;

  &lt;/div&gt;
</content>
  </entry>
</feed>
