#structureddata — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #structureddata, aggregated by home.social.
-
Declarative all the way down: Building PRONOM signatures with JSONID
by @beet_keeperPRONOM signatures are a form of declarative language, you describe the anticipated behavior in PRONOM’s regular expression syntax and tools like DROID, FIDO, and Siegfried will interpret those instructions and attempt to match them against different files to return a file format identification.
Normally, you will write PRONOM signatures by hand but doing so for file formats based on other file format building blocks can lead to inconsistencies. Bertrand Caron previously also recognized this in the XML formats that are described with PRONOM signatures on Wikidata.
XML can use single quotes ‘ (hex: 0x27) and double quotes ” (hex: 0x22) for attribute data, and so, do we make a PRONOM signature with multiple sequences anticipating the use of either?
The answer is more often than not likely to be yes, because the appearance of these values are often helpful for identifying boundaries for strings that we know must exist.
But the more file formats that we need to add to PRONOM that are based on foundational formats like XML, or JSON, or similar, the more inconsistencies will creep in, such as sequences that are looking specifically for one byte sequence over another.
The issue extends further if file formats allow data to appear at the beginning of file, or we need to account for a variable amount of white-space, or we want to start thinking about multi-byte character encoding.
We can, and in the XML issue described by myself and Caron, I think the recommendation is very much to create editorial standards for signatures for file formats based on other baseline, structured data formats like XML, JSON, YAML, and so on.
And standards are well and good, but what if tooling could help us?
For JSON this is exactly what I have tried to do in JSONID.
What does this feature look like? And what does it get us? Let’s take a look.
#declarativeProgramming #digipres #DigitalPreservation #DROID #FIDO #FileFormatIdentification #FileFormats #JSON #jsonid #JSONL #NTTW #NTTW9 #PRONOM #RDM #ResearchData #siegfried #StructuredData #structuredText #TOML #YAML -
Declarative all the way down: Building PRONOM signatures with JSONID
by @beet_keeperPRONOM signatures are a form of declarative language, you describe the anticipated behavior in PRONOM’s regular expression syntax and tools like DROID, FIDO, and Siegfried will interpret those instructions and attempt to match them against different files to return a file format identification.
Normally, you will write PRONOM signatures by hand but doing so for file formats based on other file format building blocks can lead to inconsistencies. Bertrand Caron previously also recognized this in the XML formats that are described with PRONOM signatures on Wikidata.
XML can use single quotes ‘ (hex: 0x27) and double quotes ” (hex: 0x22) for attribute data, and so, do we make a PRONOM signature with multiple sequences anticipating the use of either?
The answer is more often than not likely to be yes, because the appearance of these values are often helpful for identifying boundaries for strings that we know must exist.
But the more file formats that we need to add to PRONOM that are based on foundational formats like XML, or JSON, or similar, the more inconsistencies will creep in, such as sequences that are looking specifically for one byte sequence over another.
The issue extends further if file formats allow data to appear at the beginning of file, or we need to account for a variable amount of white-space, or we want to start thinking about multi-byte character encoding.
We can, and in the XML issue described by myself and Caron, I think the recommendation is very much to create editorial standards for signatures for file formats based on other baseline, structured data formats like XML, JSON, YAML, and so on.
And standards are well and good, but what if tooling could help us?
For JSON this is exactly what I have tried to do in JSONID.
What does this feature look like? And what does it get us? Let’s take a look.
#declarativeProgramming #digipres #DigitalPreservation #DROID #FIDO #FileFormatIdentification #FileFormats #JSON #jsonid #JSONL #NTTW #NTTW9 #PRONOM #RDM #ResearchData #siegfried #StructuredData #structuredText #TOML #YAML -
JSON-LD Explained for Personal Websites
https://hawksley.dev/blog/json-ld-explained-for-personal-websites/
#HackerNews #JSONLD #PersonalWebsites #WebDevelopment #StructuredData #TechExplained
-
JSON-LD Explained for Personal Websites
https://hawksley.dev/blog/json-ld-explained-for-personal-websites/
#HackerNews #JSONLD #PersonalWebsites #WebDevelopment #StructuredData #TechExplained
-
Von SEO zu AEO, der Kassensturz: was eine maschinenlesbare Identität wirklich bringt
Ein gutes halbes Jahr nach meinem Beitrag von SEO zu AEO der ehrliche Kassensturz: was sich gehalten hat, was naiv war und was eine maschinenlesbare, verifizierbare Identität wirklich bringt. Dazu Zahlen zu Zero-Click und KI-Antworten, jeweils mit Quelle und Grenze, und warum llms.txt kein Wundermittel ist.https://www.kernel-error.de/2026/06/12/von-seo-zu-aeo-kassensturz-maschinenlesbare-identitaet/
-
Von SEO zu AEO, der Kassensturz: was eine maschinenlesbare Identität wirklich bringt
Ein gutes halbes Jahr nach meinem Beitrag von SEO zu AEO der ehrliche Kassensturz: was sich gehalten hat, was naiv war und was eine maschinenlesbare, verifizierbare Identität wirklich bringt. Dazu Zahlen zu Zero-Click und KI-Antworten, jeweils mit Quelle und Grenze, und warum llms.txt kein Wundermittel ist.https://www.kernel-error.de/2026/06/12/von-seo-zu-aeo-kassensturz-maschinenlesbare-identitaet/
-
Here it is: Eleventy Baseline v0.1.0-next.42 is on npm.
Mostly an SEO release. Not necessarily the answer to Life, the Universe and Everything.
<baseline-head> now emits a full structured-data graph plus Open Graph, Twitter and canonical, on its own, no per-site wiring. The graph is an adapter on Joost de Valk's seo-graph-core (of Yoast fame).
Release notes: https://www.eleventy-baseline.dev/release-notes/
#StructuredData #OpenSource #BuildInPublic
#11ty #eleventy-baseline #apleasantview -
FYI: Google and Schema.org finally show how the web uses structured data: Google and Schema.org release the first public dataset on structured data term adoption, covering millions of domains in CSV and JSON formats, updated monthly. https://ppc.land/google-and-schema-org-finally-show-how-the-web-uses-structured-data/ #Google #SchemaOrg #StructuredData #DataAnalysis #WebDevelopment
-
RT @carlos_darko: Google kündigt in Zusammenarbeit mit der Schema.org-Community ein neues Repository mit Nutzungsdaten verschiedener strukturierter Datenmarkierungen im Internet an.
mehr auf Arint.info
#DataAnalytics #Google #Schemaorg #SEO #StructuredData #WebDevelopment #arint_info
-
RT @carlos_darko: Google kündigt in Zusammenarbeit mit der Schema.org-Community ein neues Repository mit Nutzungsdaten verschiedener strukturierter Datenmarkierungen im Internet an.
mehr auf Arint.info
#DataAnalytics #Google #Schemaorg #SEO #StructuredData #WebDevelopment #arint_info
-
ICYMI: Google and Schema.org finally show how the web uses structured data: Google and Schema.org release the first public dataset on structured data term adoption, covering millions of domains in CSV and JSON formats, updated monthly. https://ppc.land/google-and-schema-org-finally-show-how-the-web-uses-structured-data/ #Google #SchemaOrg #StructuredData #DataAdoption #WebDevelopment
-
Google and Schema.org finally show how the web uses structured data: Google and Schema.org release the first public dataset on structured data term adoption, covering millions of domains in CSV and JSON formats, updated monthly. https://ppc.land/google-and-schema-org-finally-show-how-the-web-uses-structured-data/ #StructuredData #SchemaOrg #Google #SEO #WebDevelopment
-
Google and Schema.org finally show how the web uses structured data: Google and Schema.org release the first public dataset on structured data term adoption, covering millions of domains in CSV and JSON formats, updated monthly. https://ppc.land/google-and-schema-org-finally-show-how-the-web-uses-structured-data/ #StructuredData #SchemaOrg #Google #SEO #WebDevelopment
-
What Is Content Engineering, and How Do You Do It?, by @lou-lin.bsky.social (@ahrefs):
https://ahrefs.com/blog/what-is-content-engineering/?ref=frontenddogma.com
-
What Is Content Engineering, and How Do You Do It?, by @lou-lin.bsky.social (@ahrefs):
https://ahrefs.com/blog/what-is-content-engineering/?ref=frontenddogma.com
-
1,885 pages had schema markup added and their AI citation rates tracked. The result? Almost no movement.
Cited pages are more likely to have JSON-LD - but that reflects technical discipline, not a lever you can pull. Authority, content clarity, and domain trust are doing the work.
#GEO #AISearch #StructuredData
https://www.digiconomy.online/insights/schema-markup-ai-citations-correlation/
-
Schema Markup ist die einzige Sprache, die Suchmaschinen ohne Interpretation verstehen. 🗣️Kein Raten, kein NLP, kein Kontext-Ableiten – du sagst Google direkt: Das ist ein Produkt, das kostet 29 €, es hat 4,7 Sterne. Wer kein strukturiertes Markup nutzt, überlässt die Kommunikation dem Zufall. Und Zufall ist keine SEO-Strategie.
-
Schema Markup ist die einzige Sprache, die Suchmaschinen ohne Interpretation verstehen. 🗣️Kein Raten, kein NLP, kein Kontext-Ableiten – du sagst Google direkt: Das ist ein Produkt, das kostet 29 €, es hat 4,7 Sterne. Wer kein strukturiertes Markup nutzt, überlässt die Kommunikation dem Zufall. Und Zufall ist keine SEO-Strategie.
-
Extra zesty post by Drew Tendero
Local search engine optimization: how to get found by nearby customers
Learn what local SEO is, why it matters, and how to optimize your business for local search. From Google Business Profile to reviews, schema markup, and local content strategies.
#SEO #SEOStrategies #SearchEngineOptimization #ContentStrategy #Google #StructuredData #SchemaMarkup
https://freshjuice.dev/blog/local-search-engine-optimization/
-
Neu im Forum:
Strukturierte Daten für Portfolio/Referenzen
https://t3forum.net/d/1100-strukturierte-daten-fuer-portfolioreferenzen
-
Neu im Forum:
Strukturierte Daten für Portfolio/Referenzen
https://t3forum.net/d/1100-strukturierte-daten-fuer-portfolioreferenzen
-
Adding AI Usage Metadata to JSON-LD Structured Data
https://rmendes.net/articles/2026/03/03/adding-ai-usage-metadata-to
-
Adding AI Usage Metadata to JSON-LD Structured Data
https://rmendes.net/articles/2026/03/03/adding-ai-usage-metadata-to
-
Technical SEO Alert: Content Types & Formats That Earn Mentions in LLMs.
FAQ Schema: 3.2x more likely to be cited in AI Overviews.
HTML Tables: 2.5x citation multiplier vs. prose.
Data Density: 3+ facts per paragraph is the new benchmark.
Architecture > Narrative for LLM retrieval. https://www.onely.com/blog/content-types-that-earn-mentions-in-llms/ -
Technical SEO Alert: Content Types & Formats That Earn Mentions in LLMs.
FAQ Schema: 3.2x more likely to be cited in AI Overviews.
HTML Tables: 2.5x citation multiplier vs. prose.
Data Density: 3+ facts per paragraph is the new benchmark.
Architecture > Narrative for LLM retrieval. https://www.onely.com/blog/content-types-that-earn-mentions-in-llms/ -
🤖 Is your website ready for AI systems?
AI-driven platforms are changing how content is discovered and processed. What matters now is not only what you publish – but how clearly your website is structured.
Our AI Readiness Workshop evaluates data structure, semantics, SEO and technical foundations – clearly prioritized and professionally assessed.
👉 Learn more and request the workshop 🔗 https://t1p.de/jqsr1
#KI #AI #AIReadiness #DigitalStrategy #WebStrategy #StructuredData
-
🤖 Is your website ready for AI systems?
AI-driven platforms are changing how content is discovered and processed. What matters now is not only what you publish – but how clearly your website is structured.
Our AI Readiness Workshop evaluates data structure, semantics, SEO and technical foundations – clearly prioritized and professionally assessed.
👉 Learn more and request the workshop 🔗 https://t1p.de/jqsr1
#KI #AI #AIReadiness #DigitalStrategy #WebStrategy #StructuredData
-
New write-up: Machine-First Search (2026 trend)
If your site isn’t machine-legible, it’s increasingly invisible. In a zero-click world, the “winner” is often the source that’s consistent, structured, and easy to verify.
https://www.sherisesstudios.com/post/the-trend-that-will-redefine-2026-machine-first-search
-
🚀 Meet the first foundation model built for tabular data—trained on a billion tables. It brings enterprise‑grade predictive intelligence to open‑source data science pipelines, now available via AWS Marketplace. Discover how this breakthrough can boost your ML workflows and unlock deeper insights from structured data. #TabularFoundationModel #BillionTableTraining #EnterpriseAI #StructuredData
🔗 https://aidailypost.com/news/fundamental-first-foundation-model-tabular-data-trained-billion-tables
-
🟦 Copilot Studio - Dataverse Structured Data
Build agents that query structured data efficiently and avoid token waste 💡💡 Fabric Data Agent
🔍 Code interpreter Python execution
⚖️ Dataverse query cacheCheck the example agent on GitHub and try it yourself.
-
🟦 Copilot Studio - Dataverse Structured Data
Build agents that query structured data efficiently and avoid token waste 💡💡 Fabric Data Agent
🔍 Code interpreter Python execution
⚖️ Dataverse query cacheCheck the example agent on GitHub and try it yourself.
-
Ist Goodreads wirklich die einzige Seite, die Bücher mittels strukturierten Daten nach Schema.org darstellt?
Ich suche das deutschsprachige Pendant, um meine Übersicht gelesener Bücher mittels WebClipper in Obsidian zu archivieren.
Wisst ihr da was?
Hier die Schema.org Docs für Bücher https://schema.org/Book
-
Ist Goodreads wirklich die einzige Seite, die Bücher mittels strukturierten Daten nach Schema.org darstellt?
Ich suche das deutschsprachige Pendant, um meine Übersicht gelesener Bücher mittels WebClipper in Obsidian zu archivieren.
Wisst ihr da was?
Hier die Schema.org Docs für Bücher https://schema.org/Book
-
SeatGeek stakes its claim in the Google agentic AI search era
https://web.brid.gy/r/https://nerds.xyz/2025/12/seatgeek-google-agentic-ai-search/
-
"The culmination of this research and foresight led my team and me to create the first Relational Foundation Model (RFM) for business data. Its purpose is to enable machines to reason directly over structured data, to understand how entities, such as customers, transactions, and products, connect. By knowing the relationships between these entities, we then enable users to make accurate predictions from those specific relationships and patterns.
Unlike LLMs, RFMs have been designed for structured relational data. RFMs are pretrained on a number of (synthetic) datasets as well as on a number of tasks over structured business data. Like LLMs, RFMs can be simply prompted to produce instant responses to a wide variety of predictive tasks over a given database, all without task-specific or database-specific training.
We wanted a system that could learn directly from how real databases are structured, and without all the usual manual setup. To make that possible, we treated each database like a graph: tables became node types, rows turned into nodes, and foreign keys linked everything together. This way, the model could actually “see” how things like customers, transactions, and products connect and change over time.
At the heart of it, the model combines a column encoder with a relational graph transformer. Every cell in a table is turned into a small numerical embedding based on what kind of data it holds, whether it’s a number, category, or a timestamp. The Transformer then looks across the graph to pull context from related tables, which helps the model adapt to new database schemas and data types.
For users to input which predictions they’d like to make, we built a simple interface called Predictive Query Language (PQL). It lets users describe what they want to predict, and the model takes care of the rest."
https://towardsdatascience.com/why-llms-arent-a-one-size-fits-all-solution-for-enterprises/
#AI #GenerativeAI #LLMs #RFMs #StructuredData #Databases #Graphs
-
How To Use (For Free) Yoast SEO Plugin FAQ Structured Data Block in WordPress? https://www.youtube.com/watch?v=RpxNFzmrCz4 🎬💡🚦❔ #FAQ #StructuredData #Guide #Yoast #SEO
-
How To Use (For Free) Yoast SEO Plugin How-To Structured Data Block in WordPress? https://www.youtube.com/watch?v=npA-WN8mcNU 🎬🚦🔍❓ #Yoast #SEO #Plugin #HowTo #StructuredData #Guide