A dedicated programme for building open, standards-based digital libraries, institutional
repositories and research archives — so that the knowledge held by Bangladesh's universities,
ministries, courts and cultural institutions is discoverable today and readable decades from now.
A digital library is not a website with PDFs on it. It is a catalogue, a preservation plan
and a set of open standards working together — and that is exactly what we build.
2Open platforms, one workflow
The Digital Library Initiative grew out of the library automation work Digital Solution
Bangladesh has delivered for the Prime Minister's Office, the National Book Center,
Jagannath University and the Chapainawabganj District Judge Court. Every one of those
projects raised the same questions — how should a Bangla-language collection be catalogued,
how do you retro-convert decades of card records, and who can still read the files in 2045?
Rather than answer those questions once per client, we made them a standing programme.
The initiative pairs production engineering with applied research and an academic advisory
team, so that what we learn on one deployment becomes reusable practice for the next
institution — documented, open and free of vendor lock-in.
Integrated library system on Koha
Institutional repository on DSpace
Bangla-aware search and cataloguing
Retro-conversion of legacy records
Preservation-grade digitisation
Staff training and written handover
Why We Do It
Mission & Vision
Two commitments guide every decision in this programme — from which metadata schema we pick
to which file format we archive in.
Our Mission
To make the recorded knowledge of Bangladesh's institutions openly discoverable, properly
catalogued and safely preserved — by deploying open-source library platforms, applying
international metadata standards, and equipping library staff to run these systems
confidently on their own.
Our Vision
A connected national knowledge network in which every university, ministry, court and
research institute in Bangladesh runs an interoperable digital library — where a single
search reaches across collections, Bangla scholarship is as machine-readable as English,
and no institution's archive is trapped inside proprietary software.
Open by default
Open-source platforms, open metadata standards and open protocols — the institution owns the system, the data and the exit route.
Standards before software
We agree the cataloguing rules and metadata profile first. The software is chosen to serve that decision, never the reverse.
Librarians lead
Professional librarians define policy and workflow; engineers implement it. Advisory oversight keeps that balance honest.
Preserved, not just stored
Checksums, archival masters, documented formats and a migration plan — storage alone is not preservation.
Capabilities
What the Initiative Delivers
Eight service lines that together take an institution from a card catalogue and a store room
to a searchable, citable, preserved digital collection.
Integrated Library System
Koha deployment covering acquisitions, cataloguing, circulation, serials, reservations, fines and branch management — configured to your service rules, not to a demo default.
Institutional Repository
DSpace repositories for theses, dissertations, journal articles, reports and datasets — with communities, collections, embargo rules and persistent handles for citation.
Digitisation & OCR
Scanning workflows for rare books, manuscripts, gazettes and judgments, producing archival masters plus searchable text through Bangla and English OCR pipelines.
Cataloguing & Retro-conversion
Original and copy cataloguing in MARC21 and Dublin Core, authority control, subject headings, and conversion of legacy card or spreadsheet records into clean bibliographic data.
Discovery & Public Catalogue
A fast, mobile-first OPAC with faceted search, Bangla transliteration handling, saved lists and a reading interface that works on a phone over a slow connection.
Interoperability & Harvesting
OAI-PMH endpoints, Z39.50 and SRU targets, SIP2 for self-service kiosks, and single sign-on against your existing identity system so the library is not an island.
Digital Preservation
Format policy, fixity checking, versioned archival storage, off-site replication and a documented migration path — designed against OAIS principles.
Training & Capacity Building
Hands-on sessions for cataloguers, circulation desk staff and system administrators, with written SOPs in Bangla and English so knowledge stays after the project closes.
Governance
Advisory Team
The initiative is guided by academics and practising librarians. They review our technical
direction, challenge our assumptions and keep the programme anchored in library science
rather than in software fashion.
Advisor
SK
Shamim Kaiser
ProfessorInstitute of Information Technology (IIT), Jahangirnagar University
Advises the initiative on its applied-computing direction — machine learning for discovery
and recommendation, OCR and document-understanding pipelines, and the responsible use of
automation in cataloguing. Also guides how the programme's research questions are framed
and how results are evaluated.
Applied AI
Machine Learning
Research Methodology
Data Analytics
Advisor
HK
Habibe Kibria Chowdhury
Deputy LibrarianJahangirnagar University
Brings the professional library perspective — collection development, cataloguing and
classification practice, reader services, and institutional repository policy. Reviews
our metadata profiles and workflow designs against how academic libraries in Bangladesh
actually operate day to day.
Library Science
Cataloguing Practice
Repository Policy
Reader Services
Seat open
Further advisors joining
Archives & Digital PreservationTo be announced
We are expanding the board with specialists in archival science, digital preservation and
open-access publishing. If you work in one of these areas and would like to contribute to
a national digital library effort, we would welcome a conversation.
Advisors meet with the engineering team at each programme milestone. They review metadata
profiles and preservation policy before a system goes live, sign off on the research agenda
for the coming cycle, and give an independent read on whether a proposed deployment genuinely
serves the institution's readers. Their guidance is advisory and non-binding, but we publish
how it shaped each decision.
Research & Innovation
Open Questions We Are Working On
Six active research tracks. Each one exists because a real deployment ran into a problem that
the off-the-shelf answer did not solve well enough for Bangla-language collections.
Bangla OCR for print & typescript
Recognition accuracy for Bangla drops sharply on aged paper, conjunct-heavy typesetting and typewritten government records. We are benchmarking pre-processing and recognition pipelines against real scanned material and publishing what actually works.
AI-assisted cataloguing
Can a model draft a MARC21 record from a title page and leave a cataloguer to verify it? We are measuring where suggestion saves time, where it introduces error, and which fields should never be automated.
Bilingual semantic search
Readers search in Bangla, in English and in romanised Bangla — often in one query. We are testing transliteration handling, stemming and embedding-based retrieval so that all three reach the same record.
Low-cost digital preservation
OAIS-aligned preservation is usually costed for well-funded archives. We are documenting a realistic tier for Bangladeshi institutions: which formats, which checksum cadence, which replication strategy, at what budget.
Cross-institution discovery
Harvesting several repositories into one search sounds simple until the metadata disagrees. We are working on a shared Dublin Core application profile and de-duplication rules for a future national union catalogue.
Usage analytics without surveillance
Libraries need circulation and access statistics; readers deserve privacy. We are designing aggregate-only analytics that inform collection decisions without building a per-reader behaviour record.
Collaborate with us
We work with departments, library schools and student researchers on these tracks — through
supervised projects, shared datasets and joint pilots on live collections. Findings, test
corpora and configuration recipes are released back to the community wherever licensing and
institutional consent allow.
Nothing here is proprietary. Every standard below is published, widely implemented and
supported by more than one piece of software — which is what makes an exit possible.
MARC21
Bibliographic and authority format for the integrated library system — the interchange language every ILS and union catalogue already speaks.
Dublin Core & MODS
Descriptive metadata for repository items, with a documented application profile so that optional fields are used consistently across collections.
OAI-PMH
Metadata harvesting endpoint, so aggregators and future national catalogues can index a collection without bespoke integration work.
Z39.50 & SRU/SRW
Federated search and copy cataloguing against external libraries, cutting original cataloguing effort for titles already described elsewhere.
SIP2 & RFID
Self-check kiosks, security gates and inventory wands connected to circulation without locking the library into one hardware vendor.
OAIS reference model
The preservation framework behind our ingest, archival storage and access design — submission, archival and dissemination packages kept distinct.
PDF/A, TIFF & JPEG 2000
Archival master and access derivative formats chosen for longevity and for tool support that is likely to outlast any single vendor.
Handle / DOI
Persistent identifiers so theses and reports stay citable even after a site is redesigned, migrated or moved to a new domain.
Platforms
Koha ILS
DSpace
PostgreSQL
Elasticsearch
Digitisation
Tesseract
ImageMagick
OCR pipelines
Batch ingest
Operations
Linux
Nightly backups
Fixity checks
Uptime monitoring
Your data, your exit
Every system we deliver can export its full bibliographic and item data in a standard format
on demand. Source code, configuration and documentation are handed over at project close.
If you ever choose to leave us, nothing about the platform makes that difficult — and we
consider that a feature, not a risk.
How a Deployment Runs
From Assessment to Handover
Five phases. Each one ends with something the institution can inspect and sign off — never a
long silence followed by a launch.
Phase 01 · Weeks 1–2
Collection & needs assessment
We survey the collection size and condition, existing records, staffing, network and space. You receive a written assessment with a realistic scope, a metadata policy proposal and an honest view of what should not be digitised.
Site survey
Records audit
Scope document
Phase 02 · Weeks 3–5
Platform setup & metadata profile
Koha and DSpace are installed, hardened and configured to your circulation rules and community structure. The cataloguing framework, controlled vocabularies and Dublin Core profile are agreed with your librarians and the advisory team before any records are loaded.
Koha config
DSpace structure
Metadata profile
Phase 03 · Weeks 4–12
Digitisation & retro-conversion
Scanning runs in batches with quality control at each one. Legacy card, register and spreadsheet records are converted to MARC21, de-duplicated and checked against authority files. Bangla and English OCR produce the searchable text layer.
Batch scanning
OCR
MARC conversion
QC sampling
Phase 04 · Weeks 10–14
Discovery, integration & go-live
The public catalogue is themed and tested on real devices, OAI-PMH and Z39.50 endpoints are opened, single sign-on and self-service hardware are connected, and the system goes live alongside a fallback until staff are comfortable.
OPAC
OAI-PMH
SSO
Soft launch
Phase 05 · Ongoing
Training, preservation & support
Role-based training for cataloguers, desk staff and administrators, with bilingual SOPs. Then scheduled backups, fixity verification, version upgrades and priority support under a maintenance agreement.
Staff training
SOPs
Fixity checks
Maintenance
Who It Serves
Built for Four Kinds of Collection
The platform is the same; the cataloguing policy, access rules and preservation priorities are
not. Each of these needs a different answer.
Universities & colleges
Course reserves, thesis and dissertation repositories, faculty publication records, and open-access policy support for departmental research output.
Government & ministries
Policy documents, gazettes, project reports and circulars — organised for officials who need the right version of a document quickly and with confidence.
Courts & legal libraries
Judgments, law reports, statutes and case files with citation-aware search and strict access control over what is public and what is not.
Archives & cultural bodies
Rare books, manuscripts, periodicals and photographs, handled with preservation-grade capture and description suited to unique, fragile material.
Questions
Digital Library FAQ
What librarians and administrators ask us most often.
Do we need both Koha and DSpace?
Not always. Koha manages the physical and licensed collection — acquisitions, circulation,
serials. DSpace holds digital objects your institution produces, such as theses, reports and
datasets. A library with no born-digital output may only need Koha; a research centre with no
lending desk may only need DSpace. We recommend after the assessment, not before it.
How well does the system handle Bangla?
Cataloguing, display and search work in Bangla throughout, and the interface can be
delivered bilingually. OCR accuracy on Bangla varies with the source: clean modern print
performs well, while aged paper, dense conjuncts and typewritten records are harder — which
is exactly why Bangla OCR is one of our active research tracks. We report expected accuracy
on a sample of your own material before committing to a digitisation plan.
What happens to our existing catalogue records?
They are migrated, not retyped. Card catalogues, registers, spreadsheets and older
software exports are converted into MARC21, de-duplicated and enriched through copy
cataloguing from external Z39.50 targets. You approve a sample batch before the full
conversion runs, and the original files are retained untouched.
Can it run on our own servers, without cloud hosting?
Yes. On-premise deployment inside your own data centre is fully supported, and for many
government and judicial clients it is the requirement. We can also host, or run a hybrid
where the catalogue is on-premise and preservation copies are replicated off-site. Backup
and fixity procedures are documented either way.
Who owns the digitised content and the metadata?
The institution does — content, metadata, configuration and source code alike. We work on
open-source foundations and open standards specifically so that ownership is real rather
than contractual. Full exports in standard formats are available at any time, not just at
the end of an engagement.
How is reader privacy protected?
Circulation history retention is configurable and defaults to the minimum the library
needs. Analytics are aggregate by design — we do not build per-reader behaviour profiles.
Access to borrower records is role-restricted and audited, and the policy is written down
and handed over rather than left to system defaults.
Can students or departments join the research work?
Yes — that is a stated aim of the initiative. We host supervised student projects, share
anonymised test corpora and run joint pilots with departments and library schools. Get in
touch with a short outline of what you would like to investigate and we will route it to the
relevant advisor.
Ready to open up your collection?
Tell us what you hold and who needs to reach it. We will come back with an assessment approach, a realistic scope and an honest view of the effort involved — no obligation.