Dhaka, Bangladesh — serving government, private & enterprise organizations

A Digital Solution Bangladesh Initiative

The Digital Library Initiative

A dedicated programme for building open, standards-based digital libraries, institutional repositories and research archives — so that the knowledge held by Bangladesh's universities, ministries, courts and cultural institutions is discoverable today and readable decades from now.

  • Koha & DSpace Open-source core
  • MARC21 & Dublin Core Standards-first metadata
  • OAI-PMH & Z39.50 Built to interoperate
  • বাংলা & English Bilingual by design
The Programme

Knowledge That Outlives Its Hardware

A digital library is not a website with PDFs on it. It is a catalogue, a preservation plan and a set of open standards working together — and that is exactly what we build.

Multi-storey university library filled with book stacks and reading galleries
2 Open platforms, one workflow

The Digital Library Initiative grew out of the library automation work Digital Solution Bangladesh has delivered for the Prime Minister's Office, the National Book Center, Jagannath University and the Chapainawabganj District Judge Court. Every one of those projects raised the same questions — how should a Bangla-language collection be catalogued, how do you retro-convert decades of card records, and who can still read the files in 2045?

Rather than answer those questions once per client, we made them a standing programme. The initiative pairs production engineering with applied research and an academic advisory team, so that what we learn on one deployment becomes reusable practice for the next institution — documented, open and free of vendor lock-in.

  • Integrated library system on Koha
  • Institutional repository on DSpace
  • Bangla-aware search and cataloguing
  • Retro-conversion of legacy records
  • Preservation-grade digitisation
  • Staff training and written handover
Why We Do It

Mission & Vision

Two commitments guide every decision in this programme — from which metadata schema we pick to which file format we archive in.

Our Mission

To make the recorded knowledge of Bangladesh's institutions openly discoverable, properly catalogued and safely preserved — by deploying open-source library platforms, applying international metadata standards, and equipping library staff to run these systems confidently on their own.

Our Vision

A connected national knowledge network in which every university, ministry, court and research institute in Bangladesh runs an interoperable digital library — where a single search reaches across collections, Bangla scholarship is as machine-readable as English, and no institution's archive is trapped inside proprietary software.

Open by default

Open-source platforms, open metadata standards and open protocols — the institution owns the system, the data and the exit route.

Standards before software

We agree the cataloguing rules and metadata profile first. The software is chosen to serve that decision, never the reverse.

Librarians lead

Professional librarians define policy and workflow; engineers implement it. Advisory oversight keeps that balance honest.

Preserved, not just stored

Checksums, archival masters, documented formats and a migration plan — storage alone is not preservation.

Capabilities

What the Initiative Delivers

Eight service lines that together take an institution from a card catalogue and a store room to a searchable, citable, preserved digital collection.

Integrated Library System

Koha deployment covering acquisitions, cataloguing, circulation, serials, reservations, fines and branch management — configured to your service rules, not to a demo default.

Institutional Repository

DSpace repositories for theses, dissertations, journal articles, reports and datasets — with communities, collections, embargo rules and persistent handles for citation.

Digitisation & OCR

Scanning workflows for rare books, manuscripts, gazettes and judgments, producing archival masters plus searchable text through Bangla and English OCR pipelines.

Cataloguing & Retro-conversion

Original and copy cataloguing in MARC21 and Dublin Core, authority control, subject headings, and conversion of legacy card or spreadsheet records into clean bibliographic data.

Discovery & Public Catalogue

A fast, mobile-first OPAC with faceted search, Bangla transliteration handling, saved lists and a reading interface that works on a phone over a slow connection.

Interoperability & Harvesting

OAI-PMH endpoints, Z39.50 and SRU targets, SIP2 for self-service kiosks, and single sign-on against your existing identity system so the library is not an island.

Digital Preservation

Format policy, fixity checking, versioned archival storage, off-site replication and a documented migration path — designed against OAIS principles.

Training & Capacity Building

Hands-on sessions for cataloguers, circulation desk staff and system administrators, with written SOPs in Bangla and English so knowledge stays after the project closes.

Governance

Advisory Team

The initiative is guided by academics and practising librarians. They review our technical direction, challenge our assumptions and keep the programme anchored in library science rather than in software fashion.

Advisor

Shamim Kaiser

Professor Institute of Information Technology (IIT), Jahangirnagar University

Advises the initiative on its applied-computing direction — machine learning for discovery and recommendation, OCR and document-understanding pipelines, and the responsible use of automation in cataloguing. Also guides how the programme's research questions are framed and how results are evaluated.

  • Applied AI
  • Machine Learning
  • Research Methodology
  • Data Analytics
Advisor

Habibe Kibria Chowdhury

Deputy Librarian Jahangirnagar University

Brings the professional library perspective — collection development, cataloguing and classification practice, reader services, and institutional repository policy. Reviews our metadata profiles and workflow designs against how academic libraries in Bangladesh actually operate day to day.

  • Library Science
  • Cataloguing Practice
  • Repository Policy
  • Reader Services
Seat open

Further advisors joining

Archives & Digital Preservation To be announced

We are expanding the board with specialists in archival science, digital preservation and open-access publishing. If you work in one of these areas and would like to contribute to a national digital library effort, we would welcome a conversation.

Express interest

How the advisory team works

Advisors meet with the engineering team at each programme milestone. They review metadata profiles and preservation policy before a system goes live, sign off on the research agenda for the coming cycle, and give an independent read on whether a proposed deployment genuinely serves the institution's readers. Their guidance is advisory and non-binding, but we publish how it shaped each decision.

Research & Innovation

Open Questions We Are Working On

Six active research tracks. Each one exists because a real deployment ran into a problem that the off-the-shelf answer did not solve well enough for Bangla-language collections.

Bangla OCR for print & typescript

Recognition accuracy for Bangla drops sharply on aged paper, conjunct-heavy typesetting and typewritten government records. We are benchmarking pre-processing and recognition pipelines against real scanned material and publishing what actually works.

AI-assisted cataloguing

Can a model draft a MARC21 record from a title page and leave a cataloguer to verify it? We are measuring where suggestion saves time, where it introduces error, and which fields should never be automated.

Bilingual semantic search

Readers search in Bangla, in English and in romanised Bangla — often in one query. We are testing transliteration handling, stemming and embedding-based retrieval so that all three reach the same record.

Low-cost digital preservation

OAIS-aligned preservation is usually costed for well-funded archives. We are documenting a realistic tier for Bangladeshi institutions: which formats, which checksum cadence, which replication strategy, at what budget.

Cross-institution discovery

Harvesting several repositories into one search sounds simple until the metadata disagrees. We are working on a shared Dublin Core application profile and de-duplication rules for a future national union catalogue.

Usage analytics without surveillance

Libraries need circulation and access statistics; readers deserve privacy. We are designing aggregate-only analytics that inform collection decisions without building a per-reader behaviour record.

Collaborate with us

We work with departments, library schools and student researchers on these tracks — through supervised projects, shared datasets and joint pilots on live collections. Findings, test corpora and configuration recipes are released back to the community wherever licensing and institutional consent allow.

Technical Foundation

Standards We Build On

Nothing here is proprietary. Every standard below is published, widely implemented and supported by more than one piece of software — which is what makes an exit possible.

MARC21
Bibliographic and authority format for the integrated library system — the interchange language every ILS and union catalogue already speaks.
Dublin Core & MODS
Descriptive metadata for repository items, with a documented application profile so that optional fields are used consistently across collections.
OAI-PMH
Metadata harvesting endpoint, so aggregators and future national catalogues can index a collection without bespoke integration work.
Z39.50 & SRU/SRW
Federated search and copy cataloguing against external libraries, cutting original cataloguing effort for titles already described elsewhere.
SIP2 & RFID
Self-check kiosks, security gates and inventory wands connected to circulation without locking the library into one hardware vendor.
OAIS reference model
The preservation framework behind our ingest, archival storage and access design — submission, archival and dissemination packages kept distinct.
PDF/A, TIFF & JPEG 2000
Archival master and access derivative formats chosen for longevity and for tool support that is likely to outlast any single vendor.
Handle / DOI
Persistent identifiers so theses and reports stay citable even after a site is redesigned, migrated or moved to a new domain.

Platforms

  • Koha ILS
  • DSpace
  • PostgreSQL
  • Elasticsearch

Digitisation

  • Tesseract
  • ImageMagick
  • OCR pipelines
  • Batch ingest

Operations

  • Linux
  • Nightly backups
  • Fixity checks
  • Uptime monitoring

Your data, your exit

Every system we deliver can export its full bibliographic and item data in a standard format on demand. Source code, configuration and documentation are handed over at project close. If you ever choose to leave us, nothing about the platform makes that difficult — and we consider that a feature, not a risk.

How a Deployment Runs

From Assessment to Handover

Five phases. Each one ends with something the institution can inspect and sign off — never a long silence followed by a launch.

  1. Phase 01 · Weeks 1–2

    Collection & needs assessment

    We survey the collection size and condition, existing records, staffing, network and space. You receive a written assessment with a realistic scope, a metadata policy proposal and an honest view of what should not be digitised.

    • Site survey
    • Records audit
    • Scope document
  2. Phase 02 · Weeks 3–5

    Platform setup & metadata profile

    Koha and DSpace are installed, hardened and configured to your circulation rules and community structure. The cataloguing framework, controlled vocabularies and Dublin Core profile are agreed with your librarians and the advisory team before any records are loaded.

    • Koha config
    • DSpace structure
    • Metadata profile
  3. Phase 03 · Weeks 4–12

    Digitisation & retro-conversion

    Scanning runs in batches with quality control at each one. Legacy card, register and spreadsheet records are converted to MARC21, de-duplicated and checked against authority files. Bangla and English OCR produce the searchable text layer.

    • Batch scanning
    • OCR
    • MARC conversion
    • QC sampling
  4. Phase 04 · Weeks 10–14

    Discovery, integration & go-live

    The public catalogue is themed and tested on real devices, OAI-PMH and Z39.50 endpoints are opened, single sign-on and self-service hardware are connected, and the system goes live alongside a fallback until staff are comfortable.

    • OPAC
    • OAI-PMH
    • SSO
    • Soft launch
  5. Phase 05 · Ongoing

    Training, preservation & support

    Role-based training for cataloguers, desk staff and administrators, with bilingual SOPs. Then scheduled backups, fixity verification, version upgrades and priority support under a maintenance agreement.

    • Staff training
    • SOPs
    • Fixity checks
    • Maintenance
Who It Serves

Built for Four Kinds of Collection

The platform is the same; the cataloguing policy, access rules and preservation priorities are not. Each of these needs a different answer.

Universities & colleges

Course reserves, thesis and dissertation repositories, faculty publication records, and open-access policy support for departmental research output.

Government & ministries

Policy documents, gazettes, project reports and circulars — organised for officials who need the right version of a document quickly and with confidence.

Courts & legal libraries

Judgments, law reports, statutes and case files with citation-aware search and strict access control over what is public and what is not.

Archives & cultural bodies

Rare books, manuscripts, periodicals and photographs, handled with preservation-grade capture and description suited to unique, fragile material.

Questions

Digital Library FAQ

What librarians and administrators ask us most often.

Do we need both Koha and DSpace?

Not always. Koha manages the physical and licensed collection — acquisitions, circulation, serials. DSpace holds digital objects your institution produces, such as theses, reports and datasets. A library with no born-digital output may only need Koha; a research centre with no lending desk may only need DSpace. We recommend after the assessment, not before it.

How well does the system handle Bangla?

Cataloguing, display and search work in Bangla throughout, and the interface can be delivered bilingually. OCR accuracy on Bangla varies with the source: clean modern print performs well, while aged paper, dense conjuncts and typewritten records are harder — which is exactly why Bangla OCR is one of our active research tracks. We report expected accuracy on a sample of your own material before committing to a digitisation plan.

What happens to our existing catalogue records?

They are migrated, not retyped. Card catalogues, registers, spreadsheets and older software exports are converted into MARC21, de-duplicated and enriched through copy cataloguing from external Z39.50 targets. You approve a sample batch before the full conversion runs, and the original files are retained untouched.

Can it run on our own servers, without cloud hosting?

Yes. On-premise deployment inside your own data centre is fully supported, and for many government and judicial clients it is the requirement. We can also host, or run a hybrid where the catalogue is on-premise and preservation copies are replicated off-site. Backup and fixity procedures are documented either way.

Who owns the digitised content and the metadata?

The institution does — content, metadata, configuration and source code alike. We work on open-source foundations and open standards specifically so that ownership is real rather than contractual. Full exports in standard formats are available at any time, not just at the end of an engagement.

How is reader privacy protected?

Circulation history retention is configurable and defaults to the minimum the library needs. Analytics are aggregate by design — we do not build per-reader behaviour profiles. Access to borrower records is role-restricted and audited, and the policy is written down and handed over rather than left to system defaults.

Can students or departments join the research work?

Yes — that is a stated aim of the initiative. We host supervised student projects, share anonymised test corpora and run joint pilots with departments and library schools. Get in touch with a short outline of what you would like to investigate and we will route it to the relevant advisor.

Ready to open up your collection?

Tell us what you hold and who needs to reach it. We will come back with an assessment approach, a realistic scope and an honest view of the effort involved — no obligation.