Generated by All in One SEO v5.0.1.1, this is an llms.txt file, used by LLMs to index the site. # Welsh Digital Grid ## Sitemaps - [XML Sitemap](https://digigrid.cymru/sitemap.xml): Contains all public & indexable URLs for this website. ## Posts - [Hello world!](https://digigrid.cymru/uncategorized/hello-world/) - Welcome to WordPress. This is your first post. Edit or delete it, then start writing! ## Pages - [Welsh Digital Grid](https://digigrid.cymru/) - Welcome to GDC-WDG, an online collection of freely available digital resources designed to support the exploration, analysis, learning, and referencing of the Welsh language. Click the images below to discover the range of resources available. The resources included on this site were developed by researchers and academics at Cardiff, Lancaster, Swansea and Bangor Universities, with - [Analyse](https://digigrid.cymru/analyse/) - FreeTxt FreeTxt supports the analysis and visualisation of free-text data in both English and Welsh (e.g. data from surveys, questionnaires, feedback fora). FreeTxt draws on some of the corpus-based utilities and methodologies from CorCenCC and ACC (below), repackaging these to enable new audiences and user-groups to analyse their own feedback data. FreeTxt is accessible to - [Links](https://digigrid.cymru/links/) - Dataset Download Link CorCenCC project's part-of-speech tagger grammar Download WNCY: WordNet Cymraeg wncy_data.sql, wncy_schema.sql, wncy.json FastText-SkipGram Welsh word embeddings Download Cross-lingual English-Welsh embeddings—trained on word2vec and FastText Label Map, Model, Tokenizer Cross-lingual English-Welsh sentiment analysis files—trained on word2vec and FastText Publication SENTimental Download - [Reference](https://digigrid.cymru/reference/) - Thesawrws The aims of the Thesaurus project, which was funded by Welsh Government, was to develop an online thesaurus of the Welsh language. Word-embeddings which are partly based on the work of CorCenCC were used to create a list of synonyms which are available freely through an online resource. Around 5000 common words were edited - [About](https://digigrid.cymru/about/) - The tools in this collection were developed in a range of collaborative UKRI and Welsh Government funded projects, in collaboration with a range of researchers, consultants and advisors representing community, industry and academic stakeholders. Details of contributors to specific tools/projects are included below. GDC-WDG is coordinated and maintained by Dawn Knight (Cardiff University), in collaboration - [Learn](https://digigrid.cymru/learn/) - Y Tiwtiadur Y Tiwtiadur is an online pedagogic toolkit which draws on data from the CorCenCC corpus. The tool includes a gap-fill or cloze exercise creator, a vocab profiler tool which colour codes words in a text according to their different frequency bands, a word identifier tool for guessing a word in context, and a - [Explore](https://digigrid.cymru/explore/) - CorCenCC CorCenCC is a freely accessible digital collection of samples of Welsh, gathered from real-life communication (a ‘corpus’). You can explore CorCenCC to find out about how people really use Welsh, for instance, how often a specific word is used, or what the most frequently used words in specific kinds of communication are. CorCenCC contains over - [Thesawrws](https://digigrid.cymru/thesawrws/) ## Categories - [Uncategorized](https://digigrid.cymru/category/uncategorized/)