Unicode text processing for the Zig programming language.
-
Updated
Oct 1, 2023
Unicode text processing for the Zig programming language.
contrib/amcheck from Postgres v11 backported to earlier Postgres versions
AQUILIGN is a multilingual alignment and collation tool for medieval texts. It uses phrase-level segmentation and contextual alignment based on BERT models, with applications in historical linguistics, philology, and historical NLP.
Open-source workbench for collating Buddhist canonical texts: punctuation diff, multi-edition collation, version lineage. CBETA/DILA integrated.
Some useful scripts for SSIS, SSRS, and SQL Server
Internationalization for Flutter and Dart on Unicode's ICU4X. Numbers, dates, plurals, lists, collation, segmentation, bidi, IDNA — same output on every platform, including web.
Forschungsdaten zur Promotionsschrift von Christian Thomas
Fast and modular collation engine with automatic variant selection
Correct Turkish casing, folding, slugs and sorting. Handles the dotted/dotless i without needing ICU. Zero dependencies.
Collate Greek New Testament manuscripts
[DEPRECATED] Implementation of an Ethereum sharding collation
A command-line tool for organizing and collating text files in writing projects.
A Zig native Unicode (v17) engine: UTF-8/16/32, normalization, segmentation, unicode properties, bidi, collation (DUCET)
Typography-first collation reader for comparing two witnesses of a literary text. Word-level variants, moved/split/merged passage detection, verse-aware comparison, synoptic and unified reading, and TEI P5 export.
Add a description, image, and links to the collation topic page so that developers can more easily learn about it.
To associate your repository with the collation topic, visit your repo's landing page and select "manage topics."