AI/LLM-based tooling and best practices for semantic data pipelines #24
Replies: 7 comments 1 reply
List of initiatives/platforms in the room:
LLM-y stuff:
Non-LLMy stuff: I. Brainstorming ELN interop use cases -- ELNFileFormat breakout? (Edgar the ELN) Potential deliverables:
Breakouts for the week:
|
|
All the non-LLMy stuff is here: |
|
All LLM prompts, codes etc are in the "herbert" repo: |
|
Summary 21.+22.10.2025:
|
|
Summary Martin's work 23.10.2025
|
|
Summary of work on rdm-agent Simple langchain-based scaffold for "I have some data in a directory that I would like to be reviewed", with the model prompted to consider aspects of FAIR principles, dark/missing data or even missing fields from a schema (e.g., a processing step without including the processing parameters). The agent has "tools" (Python functions) available to it that can perform subtasks, e.g., if a schema is provided, run a schema validation of the data provided. Next steps:
|
|
Summary of work on Herbert 24/10/25 A pull request has been made for the MADICES/herbert repo containing initial code for the 'agent' Herbert. Hopefully this pull request will be merged into Key features of this initial version are as follows:
Possible next steps:
Current system prompt for Herbert: |
Uh oh!
There was an error while loading. Please reload this page.
Working group notes building on #3
All reactions