4 ms·
Your HTML already has semantic meta elements like author and description you should be populating with info like that: https://developer.mozilla.org/en-US/docs/
by qbasic_forever 3y ago
Your HTML already has semantic meta elements like author and description you should be populating with info like that: https://developer.mozilla.org/en-US/docs/Learn/HTML/Introduction_to_HTML/The_head_metadata_in_HTML https://developer.mozilla.org/en-US/docs/Learn/HTML/Introduc...
- techaqua 3y agoand also opengraph meta tags https://ogp.me/ https://ogp.me/
- doodlesdev 3y agoAnd also schema.org: https://schema.org/ https://schema.org/
- westurner 3y agoThing > CreativeWork > WebSite https://schema.org/WebSite https://schema.org/WebSite ... scroll down to "Examples" and click the "JSON-LD" and/or "RDFa" tabs. (And if there isn't an example then go to the schema.org/ URL of a superClassOf (rdfs:subClassOf) of the rdfs:Class or rdfs:Property; there are many markup examples for CreativeWork and subtypes). httpS://schema.org/license Also: https://news.ycombinator.com/item?id=35891631 https://news.ycombinator.com/item?id=35891631 extruct is one way to parse linked data from HTML pages: https://github.com/scrapinghub/extruct https://github.com/scrapinghub/extruct
- burnte 3y agoHow do I add a semantic definition in an HTML tag to a JPEG, or MP4, or WAV, or any non HTML format? HTML tags fix HTML, not other formats.
- akira2501 3y agoWhat would the difference in semantic notation between an unstructured "ai.txt" and the "alt" attribute actually be? If you want the tags to be served with the context outside of HTML, you can always use HTML header attributes.
- cornstalks 3y agoJPEG has EXIF, MP3 has ID3 tags, MP4 has ilst, MKV has Tags, etc. We don't need xkcd/927 for these other formats that already have standard metadata mechanisms.
- RamblingCTO 3y agoYeah, and if you make it that complex to extract consent you won't get any. Think one step ahead maybe. One switch, one thing to parse.
- qbasic_forever 3y agoRobots.txt is where you tell crawlers (AI or otherwise) what should and shouldn't be read on your site. Metadata like in tags, HTML meta tags, etc. is where you describe the content so meaning can be extracted from it by machines and automated processing.
- cornstalks 3y ago1. OP said “what it is about, when was it published, the author, etc.” That’s what these mechanisms already cover. Consent is an interesting possibility that I’ll admit something like ai.txt might be better for, but my post was largely focused on the OP. 2. These are all complex formats. If you want to ingest and process them then you already have to build all the hard parts. Getting the metadata out is dead simple compared to parsing, decoding, and then processing an image, for example.
- qbasic_forever 3y agoIf you're describing an object on the page, like an image or video, you want a label element linked to it by id and likely an aria-label attribute on the object. (screen readers and such will look for this in particular). For an image you want an alt text attribute as a description too.