Currently, there's no real content negotiation for images in srcset or
source tags, only something based on the optional type attribute.
Part of the conversion process consists of converting picture elements
to img with an srcset attribute.
Before archival, every img[srcset] attribute is reconstructed
in order to sort the list and have the biggest candidates first (and
discard images that will be way too big)
The archiver will then try to download every image in the given order
and, once successful, set an src attribute to the image (with the winner)
and remove the srcset.
Finally, this adds 2 new steps in the CleanDomProcessor:
- remove empty id attributes
- remove img with no src and not srcset attribute
- golangci-lint configuration
- gofumpt on all go files
- added package comments
- added missing comments on exported functions
- do not check for bodyclose on http testing responses.
- check for errors on triggered tasks
- check for errors in acls
- check for errors during policy loading
- check for errors in password recovery process
- better error checking in pkg/extract/contentscripts
- nolint rules for errcheck when it's not needed
(mostly defer calls for file reader closing)
This was a mistake to use go.work since it might be needed on a
specific dev setup (with different "use" directives).
It results in a codebase that really matches the main readeck
module fqdn.