| Commit message (Collapse) | Author | Age | Files | Lines |
| |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| |
Five new datasets with provenance in SOURCES files:
* UWG: 3. UWGAendG (BGBl. 2026 I Nr. 43) in the new digital BGBl
format, plus drafts.
* GEG: the 2023 "Heizungsgesetz" (BGBl. 2023 I Nr. 280, new format)
and the pending GModG 2026 drafts.
* AGG: 2. AGGAendG drafts (RefE/RegE/BT-Drs 21/6178) -- the base XML
predates the bill, so this dataset produces real diffs -- and the
official BMJV synopsis as ground truth.
* ProdHaftG: product-liability modernization drafts (Artikel 1 is a
replacement act, Artikel 2 amends the old law), with official
synopsis.
* BayJG: Bavarian hunting-law amendment (GVBl. 2026 S. 113) -- state
law is not yet supported by the tool; kept for future work.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Change-Id: I8c33b6a19178f73cece9c9da99c713009016e578
|
| |
|
|
|
|
|
|
|
|
|
| |
The Infektionsschutzgesetz as gii-norm XML, EPUB, and PDF (consolidated
as of 2020-11-20, i.e. already including the Art. 1 changes below), the
Drittes Bevoelkerungsschutzgesetz (BGBl. I 2020 S. 2397) as published,
and the Bundestag/Bundesrat drafting documents, with provenance URLs in
SOURCES. Used by the end-to-end smoke test.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Change-Id: I1221734ea3e3dab705151421774151438be1b337
|
| |
|
|
|
|
|
|
|
|
|
|
|
|
| |
Bump all dependencies and plugins to current versions. Replace the
tika-parsers bundle with a direct PDFBox dependency (we need
PDFTextStripper control), add java-diff-utils for the synopsis diff,
and add JUnit 5 with AssertJ for testing. Drop unused dependencies
(tess4j, Lanterna, term4j, sqlite-jdbc, imageio codecs, annotation
libraries) in favour of jspecify. Regenerate the Maven wrapper with
the official plugin (Takari is dead) and remove the vestigial Ant
wrapper and Tika configuration.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Change-Id: I0e1813b3967027cbb0bb978591d345ce138fa0a0
|
|
|
With the new configuration, Tika can now extract text from PDFs and
XML documents.
Also configures logging for the application.
Change-Id: I7a89c2b232ed4e220665dd335a5f5a0cc3ef2994
|