Explore projects
-
Thèse Guillaume Bernard / Développement / from events to documents / database_infrastructure_text_mining
GNU General Public License v3.0 or laterTextual Search Engine Infrastructure based on ElasticSearch (https://www.elastic.co/fr/elasticsearch/) and Lucene (https://lucene.apache.org/). Includes the import scripts to load datasets into the index.
archived 0Updated -
Thèse Guillaume Bernard / Jeux de données / dataset_manipulation_tools / synthesise_ocr_and_segmentation_errors_in_texts
GNU General Public License v3.0 or laterThis software enables to damage texts written in any natural language by applying OCR degradation (phantom characters, character degradation, etc.) and by over-segmenting texts (this means splitting regularly the texts in equal parts).
This is useful to reproduce common errors found in historical documents when historical data is missing.
archived 0Updated -
Updated
-
Updated
-
Updated
-
Updated
-
Updated
-
Updated
-
S4 Structure de Données TP1 Immutable Linked List
Updated -
Updated
-
-
Updated
-
Visualisation du registre des traitements / Application web de visualisation du registre légal des traitements
CeCILL-B Free Software License AgreementProjet Open Source de visualisation interactive du registre des traitements de l'agglomération de La Rochelle.
Updated -
-
Updated