Blog

Read blog posts written by members of our team about the Language Research domain.

2026

LDaCA at the Zenadth Kes Language Symposium 2026

LDaCA at the Zenadth Kes Language Symposium 2026

Read about LDaCA team members Alistair Harvey and Ben Foley's presentation at the Zenadth Kes Language Symposium in Thursday Island, focusing on community engagement and community archiving projects.

Interview with Al Harvey

Interview with Al Harvey

Meet our newest Industry Fellow, Al Harvey, and learn about his work with community archives.

Inside the GDRF2026: Teresa Chan on her experience so far

Inside the GDRF2026: Teresa Chan on her experience so far

What is it actually like to take part in LDaCA's Graduate Digital Research Fellowship (GDRF)? We put the question to Teresa Chan, LDaCA’s Senior Research Project Officer in Research Support and Training and a 2026 GDRF participant.

Graduate Digital Research Fellowship — 2025

Graduate Digital Research Fellowship — 2025

A blog post about the 2025 Graduate Digital Research Fellowship cohort, co-ordinated by Sam Hames (Research Analytics Lead) and Simon Musgrave (Research Support and Training Lead).

Sharing the Australian Slang Survey data collection

Sharing the Australian Slang Survey data collection

The Australian Slang Survey asked participants about expressions they thought of as typically Australian. Language Technology Analyst Rosanna Smith and Research Support and Training Lead Simon Musgrave worked to make this dataset FAIR compliant and publish it on the LDaCA Data Portal.

2025

AI and Language Data: Setting the Groundwork for 2026

AI and Language Data: Setting the Groundwork for 2026

LDaCA is assembling a working party of those in our network using AI in their research. We intend to release advice covering particular AI systems, to be used as a guide for creating and working with language data using AI.

Analyse image collections with the Image Dataset Explorer

Analyse image collections with the Image Dataset Explorer

Learn more about how Research Analytics Lead Sam Hames and University of Queensland student Jasper Chong developed the Image Dataset Explorer, a tool for making sense of large image collections. Their work revealed insights that can also be applied to working with text.

Implementing PILARS

Implementing PILARS

Discover how we have been working towards implementing PILARS — the Protocols for Implementing Long-Term Archival Repository Services — by adopting open standards, building clear governance mechanisms and designing infrastructure that communities can trust and control.

Corpus spotlight: Mitchell and Delbridge

Corpus spotlight: Mitchell and Delbridge

Read about one of the first collections to enter our data portal — the Mitchell and Delbridge speech of Australian adolescents corpus (1959–1960) — and its impact on Australian linguistic research.

LDaCA Technical Architecture update 2025

LDaCA Technical Architecture update 2025

Refresh your knowledge of the principles behind our technical architecture and discover some recent developments, including how we are harmonising the open source tools used across our network of collaborators.

Putting data to work — 2

Putting data to work — 2

Learn more about the challenges of working with an unwieldy data source — the Australian Federal Hansard — from Simon Musgrave. In the blog post, Simon unpacks approaches to issues of access and scale allowing for the effective use of this data source.

Interview with Nick Thieberger

Interview with Nick Thieberger

Chief Investigator Nick Thieberger (University of Melbourne) discusses storage and archival practice at PARADISEC, dialogic archives and the future of long term storage.

Team member tip: Mastering metadata

Team member tip: Mastering metadata

Data Migration Developer Mark Raadgever draws on his extensive experience with data migration to outline some important metadata principles. Consistency is key!

Reflecting on the Darwin digital languages collections workshop

Reflecting on the Darwin digital languages collections workshop

From 24–28 March, the LDaCA team hosted the Darwin digital languages collections workshop. This event brought together organisations and individuals from the Top End to exchange ideas and insights about Indigenous language collections and explore collaborative opportunities.

Putting data to work

Putting data to work

Research Support and Training Lead Simon Musgrave discusses how he used four collections from the LDaCA portal to strengthen his argument in a recent publication that there is a tradition of Australian writers inventing expressions which they treat as being part of Australian slang.

CAREful FAIRness principles for Indigenous Data Governance

CAREful FAIRness principles for Indigenous Data Governance

Robert McLellan, Senior Program Manager for LDaCA, and Jenny Fewster, Director, HASS and Indigenous Research Data Commons for the Australian Research Data Commons, draw upon their collaborative work on CAREful FAIRness and Principles for Indigenous Data Governance.

Indigenous data governance: A discussion

Indigenous data governance: A discussion

Read a summary of our learnings from the LDaCA-hosted Indigenous Data Governance panel discussion, which brought together three speakers — Lesley Acres (UQ Library), Dr Rose Barrowcliffe (Macquarie University and LDaCA), and Robert McLellan (UQ and LDaCA).

2024

Interview with Jane Simpson

Interview with Jane Simpson

Jane Simpson, Professor Emerita at the Australian National University, discusses storing language data, obstacles to its long-term storage, her interest in making dictionaries accessible and how researchers at the end of their careers should manage data.

Graduate Digital Research Fellowship — 2024

Graduate Digital Research Fellowship — 2024

A blog post about the 2024 Graduate Digital Research Fellowship cohort, co-ordinated by Sam Hames (Research Analytics Lead) and Simon Musgrave (Research Support and Training Lead).

Interview with Kalin Stefanov

Interview with Kalin Stefanov

An interview with Kalin Stefanov (ARC DECRA Fellow at Monash University) about his work on an Australian Sign Language (Auslan) project.

Graduate Digital Research Fellowship — 2023

Graduate Digital Research Fellowship — 2023

A blog post about the 2023 Graduate Digital Research Fellowship cohort, co-ordinated by Sam Hames (Research Analytics Lead) and Simon Musgrave (Research Support and Training Lead).

2023

Interview with our Chief Investigators - Part 2

Interview with our Chief Investigators - Part 2

An interview with some former and current LDaCA Chief Investigators. This blog post features Catherine Travis (Australian National University), Monika Bednarek (University of Sydney) and former Chief Investigator Nicholas Evans (Australian National University).

Introducing the Keyword Analysis tool

Introducing the Keyword Analysis tool

The Keyword Analysis tool is a Jupyter notebook containing code that was developed by the Sydney Informatics Hub (SIH) in collaboration with the Sydney Corpus Lab.

Interview with our Chief Investigators - Part 1

Interview with our Chief Investigators - Part 1

An interview with some former and current LDaCA Chief Investigators. This post features Martin Schweinberger (University of Queensland), Nick Thieberger (University of Melbourne) and former Chief Investigator Louisa Willoughby (Monash University).

2022

Discursis

Discursis

Discursis is communication analytics technology for analysing text-based communication data, participant interactions, topics and inter-speaker relationships.