visual-parser 2.0.2__tar.gz → 2.0.4__tar.gz
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- {visual_parser-2.0.2 → visual_parser-2.0.4}/PKG-INFO +22 -17
- {visual_parser-2.0.2 → visual_parser-2.0.4}/README.md +157 -152
- {visual_parser-2.0.2 → visual_parser-2.0.4}/pyproject.toml +1 -1
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/__init__.py +1 -1
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser.egg-info/PKG-INFO +22 -17
- {visual_parser-2.0.2 → visual_parser-2.0.4}/setup.cfg +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/__main__.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/cli.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/cli_main.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/config.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/figure_describer.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/jsonl_writer.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/metadata_extractor.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/nougat_engine.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/pdf_tracker.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/pipeline.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/prompts.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/text_extractor.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser/vision_llm.py +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser.egg-info/SOURCES.txt +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser.egg-info/dependency_links.txt +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser.egg-info/entry_points.txt +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser.egg-info/requires.txt +0 -0
- {visual_parser-2.0.2 → visual_parser-2.0.4}/visual_parser.egg-info/top_level.txt +0 -0
|
@@ -1,6 +1,6 @@
|
|
|
1
1
|
Metadata-Version: 2.4
|
|
2
2
|
Name: visual-parser
|
|
3
|
-
Version: 2.0.
|
|
3
|
+
Version: 2.0.4
|
|
4
4
|
Summary: Standalone Visual-RAG PDF Parser - text extraction and Vision-LLM figure descriptions to JSONL
|
|
5
5
|
Author: Zavier N. Ndum
|
|
6
6
|
Author-email: zavier.ndum@tamu.edu
|
|
@@ -43,7 +43,6 @@ Requires-Dist: mypy; extra == "dev"
|
|
|
43
43
|
|
|
44
44
|
# visual-parser (Standalone Visual-RAG PDF Ingestion)
|
|
45
45
|
|
|
46
|
-
<!--  -->
|
|
47
46
|

|
|
48
47
|
<!--  -->
|
|
49
48
|
|
|
@@ -75,15 +74,16 @@ Prebuilt images are on **[zev94/radiant-llm](https://hub.docker.com/r/zev94/radi
|
|
|
75
74
|
|
|
76
75
|
| Tag | Description |
|
|
77
76
|
|-----|-------------|
|
|
78
|
-
| `visual-parser-
|
|
79
|
-
| `visual-parser-
|
|
77
|
+
| `visual-parser-latest` | Always latest build (rolling) |
|
|
78
|
+
| `visual-parser-2.0.0` | Pinned release (Apache 2.0 release) |
|
|
79
|
+
| `visual-parser-1.0` | Legacy — v1.0.0, stale |
|
|
80
80
|
|
|
81
81
|
### 1) Install Docker
|
|
82
82
|
- Docker Desktop (Windows/macOS) or Docker Engine (Linux)
|
|
83
83
|
|
|
84
84
|
### 2) Pull the image
|
|
85
85
|
```bash
|
|
86
|
-
docker pull zev94/radiant-llm:visual-parser-
|
|
86
|
+
docker pull zev94/radiant-llm:visual-parser-latest
|
|
87
87
|
```
|
|
88
88
|
|
|
89
89
|
### 3) Run (input + output on the same mounted folder)
|
|
@@ -91,7 +91,7 @@ Windows PowerShell:
|
|
|
91
91
|
```powershell
|
|
92
92
|
docker run --rm --env-file .env `
|
|
93
93
|
-v "C:\path\to\pdfs:/data" `
|
|
94
|
-
zev94/radiant-llm:visual-parser-
|
|
94
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
95
95
|
--input-dir /data --output-dir /data
|
|
96
96
|
```
|
|
97
97
|
|
|
@@ -99,7 +99,7 @@ Linux / WSL:
|
|
|
99
99
|
```bash
|
|
100
100
|
docker run --rm --env-file .env \
|
|
101
101
|
-v "/path/to/pdfs:/data" \
|
|
102
|
-
zev94/radiant-llm:visual-parser-
|
|
102
|
+
zev94/radiant-llm:visual-parser-latest \
|
|
103
103
|
--input-dir /data --output-dir /data
|
|
104
104
|
```
|
|
105
105
|
|
|
@@ -109,7 +109,7 @@ Windows PowerShell:
|
|
|
109
109
|
docker run --rm --env-file .env `
|
|
110
110
|
-v "C:\path\to\pdfs:/data" `
|
|
111
111
|
-v "C:\path\to\out:/out" `
|
|
112
|
-
zev94/radiant-llm:visual-parser-
|
|
112
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
113
113
|
--input-dir /data --output-dir /out
|
|
114
114
|
```
|
|
115
115
|
|
|
@@ -122,11 +122,11 @@ docker images # use the tag printed by Docker
|
|
|
122
122
|
|
|
123
123
|
### Model overrides (optional)
|
|
124
124
|
|
|
125
|
-
Default vision model is **GPT-5.
|
|
125
|
+
Default vision model is **GPT-5.4** when using `--vision-provider gpt`. Override on the command line:
|
|
126
126
|
|
|
127
127
|
```powershell
|
|
128
128
|
docker run --rm --env-file .env -v "C:\path\to\pdfs:/data" `
|
|
129
|
-
zev94/radiant-llm:visual-parser-
|
|
129
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
130
130
|
--input-dir /data --output-dir /data --vision-model gpt-5.4
|
|
131
131
|
```
|
|
132
132
|
|
|
@@ -142,7 +142,7 @@ python visual-parser.py --input-dir "C:\path\to\pdfs"
|
|
|
142
142
|
After pulling the image, run:
|
|
143
143
|
|
|
144
144
|
```bash
|
|
145
|
-
docker run --rm zev94/radiant-llm:visual-parser-
|
|
145
|
+
docker run --rm zev94/radiant-llm:visual-parser-latest --help
|
|
146
146
|
```
|
|
147
147
|
|
|
148
148
|
For copy-paste **Docker** examples (vision presets, text modes, workers, rebuild), see [`docker-usage-examples.md`](docker-usage-examples.md).
|
|
@@ -173,23 +173,28 @@ Performance / misc:
|
|
|
173
173
|
|
|
174
174
|
## Citation
|
|
175
175
|
|
|
176
|
-
If you use RADIANT-LLM or the accompanying evaluation materials, please cite the
|
|
176
|
+
If you use RADIANT-LLM or the accompanying evaluation materials, please cite the journal article:
|
|
177
177
|
|
|
178
178
|
```bibtex
|
|
179
|
-
@article{
|
|
180
|
-
title={
|
|
179
|
+
@article{ndum2026retrieval,
|
|
180
|
+
title={A retrieval-augmented, domain-intelligent agentic framework for reliable decision support in safety-critical nuclear engineering},
|
|
181
181
|
author={Ndum, Zavier Ndum and Tao, Jian and Ford, John and Yim, Mansung and Liu, Yang},
|
|
182
|
-
journal={
|
|
183
|
-
|
|
182
|
+
journal={Reliability Engineering \& System Safety},
|
|
183
|
+
pages={113057},
|
|
184
|
+
year={2026},
|
|
185
|
+
publisher={Elsevier}
|
|
184
186
|
}
|
|
185
187
|
```
|
|
186
188
|
|
|
189
|
+
Journal: *Reliability Engineering & System Safety* (2026), article 113057
|
|
187
190
|
Preprint: https://arxiv.org/abs/2604.22755
|
|
188
191
|
|
|
189
192
|
---
|
|
190
193
|
|
|
191
194
|
## License
|
|
192
195
|
|
|
193
|
-
|
|
196
|
+
Copyright 2026 Zavier N. Ndum
|
|
197
|
+
|
|
198
|
+
This project is licensed under the Apache License 2.0. See the [LICENSE](https://github.com/SmartLabNuclear/RADIANT_LLM/blob/main/LICENSE) file in the RADIANT_LLM repository for the full license text.
|
|
194
199
|
|
|
195
200
|
|
|
@@ -1,152 +1,157 @@
|
|
|
1
|
-
# visual-parser (Standalone Visual-RAG PDF Ingestion)
|
|
2
|
-
|
|
3
|
-
|
|
4
|
-
.
|
|
106
|
-
|
|
107
|
-
Paths:
|
|
108
|
-
- `--input-dir` / `-i` (required)
|
|
109
|
-
- `--output-dir` / `-o` (default: same as input)
|
|
110
|
-
|
|
111
|
-
Text extraction:
|
|
112
|
-
- `--text-mode nougat|lightweight` (default: `nougat`)
|
|
113
|
-
- `--nougat-model facebook/nougat-small`
|
|
114
|
-
- `--chunk-size 500`
|
|
115
|
-
- `--chunk-overlap 100`
|
|
116
|
-
|
|
117
|
-
Vision LLM:
|
|
118
|
-
- `--vision-provider gpt|gemini` (default: `gpt`)
|
|
119
|
-
- `--vision-model gpt-5.2` (or `gpt-4o`, `gemini-2.5-flash`, etc.)
|
|
120
|
-
- `--vision-detail low|high|auto`
|
|
121
|
-
- `--reasoning-effort none|low|medium|high|xhigh`
|
|
122
|
-
- `--metadata-pages 2`
|
|
123
|
-
|
|
124
|
-
Performance / misc:
|
|
125
|
-
- `--max-workers 4`
|
|
126
|
-
- `--rebuild` (reprocess everything; ignore `04_processed_pdfs.txt`)
|
|
127
|
-
- `--log-level DEBUG|INFO|WARNING|ERROR`
|
|
128
|
-
|
|
129
|
-
---
|
|
130
|
-
|
|
131
|
-
## Citation
|
|
132
|
-
|
|
133
|
-
If you use RADIANT-LLM or the accompanying evaluation materials, please cite the
|
|
134
|
-
|
|
135
|
-
```bibtex
|
|
136
|
-
@article{
|
|
137
|
-
title={
|
|
138
|
-
author={Ndum, Zavier Ndum and Tao, Jian and Ford, John and Yim, Mansung and Liu, Yang},
|
|
139
|
-
journal={
|
|
140
|
-
|
|
141
|
-
}
|
|
142
|
-
|
|
143
|
-
|
|
144
|
-
|
|
145
|
-
|
|
146
|
-
|
|
147
|
-
|
|
148
|
-
|
|
149
|
-
|
|
150
|
-
|
|
151
|
-
|
|
152
|
-
|
|
1
|
+
# visual-parser (Standalone Visual-RAG PDF Ingestion)
|
|
2
|
+
|
|
3
|
+

|
|
4
|
+
<!--  -->
|
|
5
|
+
|
|
6
|
+
`visual-parser` is a standalone document-ingestion tool that converts PDFs into a multi-modal JSONL knowledge base (text chunks + figure descriptions + metadata). The intended workflow is:
|
|
7
|
+
|
|
8
|
+
1) Run `visual-parser` on curated PDFs to generate JSONL KB files.
|
|
9
|
+
2) Run RADIANT-LLM Visual-RAG for QA over the generated KB.
|
|
10
|
+
|
|
11
|
+
## Outputs (JSONL KB)
|
|
12
|
+
|
|
13
|
+
By default, the pipeline writes:
|
|
14
|
+
- `01_chunks_kb.jsonl`: chunked text extracted from PDFs (Nougat by default).
|
|
15
|
+
- `02_visuals_kb.jsonl`: figure/page visual descriptions (Vision LLM).
|
|
16
|
+
- `03_metadata_kb.jsonl`: document metadata rows (title/author/etc.).
|
|
17
|
+
- `04_processed_pdfs.txt`: a tracker so re-runs only process new PDFs (unless `--rebuild`).
|
|
18
|
+
|
|
19
|
+
## API keys (`.env`)
|
|
20
|
+
|
|
21
|
+
Provide at least one provider:
|
|
22
|
+
- `OPENAI_API_KEY` (OpenAI)
|
|
23
|
+
- `GEMINI_API_KEY` (Gemini)
|
|
24
|
+
|
|
25
|
+
Optional:
|
|
26
|
+
- `HF_TOKEN` (if you use gated Hugging Face models)
|
|
27
|
+
|
|
28
|
+
## Run with Docker (Docker Hub)
|
|
29
|
+
|
|
30
|
+
Prebuilt images are on **[zev94/radiant-llm](https://hub.docker.com/r/zev94/radiant-llm)** under the **visual-parser** tags:
|
|
31
|
+
|
|
32
|
+
| Tag | Description |
|
|
33
|
+
|-----|-------------|
|
|
34
|
+
| `visual-parser-latest` | Always latest build (rolling) |
|
|
35
|
+
| `visual-parser-2.0.0` | Pinned release (Apache 2.0 release) |
|
|
36
|
+
| `visual-parser-1.0` | Legacy — v1.0.0, stale |
|
|
37
|
+
|
|
38
|
+
### 1) Install Docker
|
|
39
|
+
- Docker Desktop (Windows/macOS) or Docker Engine (Linux)
|
|
40
|
+
|
|
41
|
+
### 2) Pull the image
|
|
42
|
+
```bash
|
|
43
|
+
docker pull zev94/radiant-llm:visual-parser-latest
|
|
44
|
+
```
|
|
45
|
+
|
|
46
|
+
### 3) Run (input + output on the same mounted folder)
|
|
47
|
+
Windows PowerShell:
|
|
48
|
+
```powershell
|
|
49
|
+
docker run --rm --env-file .env `
|
|
50
|
+
-v "C:\path\to\pdfs:/data" `
|
|
51
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
52
|
+
--input-dir /data --output-dir /data
|
|
53
|
+
```
|
|
54
|
+
|
|
55
|
+
Linux / WSL:
|
|
56
|
+
```bash
|
|
57
|
+
docker run --rm --env-file .env \
|
|
58
|
+
-v "/path/to/pdfs:/data" \
|
|
59
|
+
zev94/radiant-llm:visual-parser-latest \
|
|
60
|
+
--input-dir /data --output-dir /data
|
|
61
|
+
```
|
|
62
|
+
|
|
63
|
+
### 4) Run (separate output directory)
|
|
64
|
+
Windows PowerShell:
|
|
65
|
+
```powershell
|
|
66
|
+
docker run --rm --env-file .env `
|
|
67
|
+
-v "C:\path\to\pdfs:/data" `
|
|
68
|
+
-v "C:\path\to\out:/out" `
|
|
69
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
70
|
+
--input-dir /data --output-dir /out
|
|
71
|
+
```
|
|
72
|
+
|
|
73
|
+
### Offline install (legacy `.tar`)
|
|
74
|
+
|
|
75
|
+
```powershell
|
|
76
|
+
docker load -i .\visual-parser_0.1.0.tar
|
|
77
|
+
docker images # use the tag printed by Docker
|
|
78
|
+
```
|
|
79
|
+
|
|
80
|
+
### Model overrides (optional)
|
|
81
|
+
|
|
82
|
+
Default vision model is **GPT-5.4** when using `--vision-provider gpt`. Override on the command line:
|
|
83
|
+
|
|
84
|
+
```powershell
|
|
85
|
+
docker run --rm --env-file .env -v "C:\path\to\pdfs:/data" `
|
|
86
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
87
|
+
--input-dir /data --output-dir /data --vision-model gpt-5.4
|
|
88
|
+
```
|
|
89
|
+
|
|
90
|
+
<!-- ## Run from source (Python)
|
|
91
|
+
|
|
92
|
+
From `codebase/Visual-Parser/`:
|
|
93
|
+
```powershell
|
|
94
|
+
python visual-parser.py --input-dir "C:\path\to\pdfs"
|
|
95
|
+
``` -->
|
|
96
|
+
|
|
97
|
+
## Common configuration flags
|
|
98
|
+
|
|
99
|
+
After pulling the image, run:
|
|
100
|
+
|
|
101
|
+
```bash
|
|
102
|
+
docker run --rm zev94/radiant-llm:visual-parser-latest --help
|
|
103
|
+
```
|
|
104
|
+
|
|
105
|
+
For copy-paste **Docker** examples (vision presets, text modes, workers, rebuild), see [`docker-usage-examples.md`](docker-usage-examples.md).
|
|
106
|
+
|
|
107
|
+
Paths:
|
|
108
|
+
- `--input-dir` / `-i` (required)
|
|
109
|
+
- `--output-dir` / `-o` (default: same as input)
|
|
110
|
+
|
|
111
|
+
Text extraction:
|
|
112
|
+
- `--text-mode nougat|lightweight` (default: `nougat`)
|
|
113
|
+
- `--nougat-model facebook/nougat-small`
|
|
114
|
+
- `--chunk-size 500`
|
|
115
|
+
- `--chunk-overlap 100`
|
|
116
|
+
|
|
117
|
+
Vision LLM:
|
|
118
|
+
- `--vision-provider gpt|gemini` (default: `gpt`)
|
|
119
|
+
- `--vision-model gpt-5.2` (or `gpt-4o`, `gemini-2.5-flash`, etc.)
|
|
120
|
+
- `--vision-detail low|high|auto`
|
|
121
|
+
- `--reasoning-effort none|low|medium|high|xhigh`
|
|
122
|
+
- `--metadata-pages 2`
|
|
123
|
+
|
|
124
|
+
Performance / misc:
|
|
125
|
+
- `--max-workers 4`
|
|
126
|
+
- `--rebuild` (reprocess everything; ignore `04_processed_pdfs.txt`)
|
|
127
|
+
- `--log-level DEBUG|INFO|WARNING|ERROR`
|
|
128
|
+
|
|
129
|
+
---
|
|
130
|
+
|
|
131
|
+
## Citation
|
|
132
|
+
|
|
133
|
+
If you use RADIANT-LLM or the accompanying evaluation materials, please cite the journal article:
|
|
134
|
+
|
|
135
|
+
```bibtex
|
|
136
|
+
@article{ndum2026retrieval,
|
|
137
|
+
title={A retrieval-augmented, domain-intelligent agentic framework for reliable decision support in safety-critical nuclear engineering},
|
|
138
|
+
author={Ndum, Zavier Ndum and Tao, Jian and Ford, John and Yim, Mansung and Liu, Yang},
|
|
139
|
+
journal={Reliability Engineering \& System Safety},
|
|
140
|
+
pages={113057},
|
|
141
|
+
year={2026},
|
|
142
|
+
publisher={Elsevier}
|
|
143
|
+
}
|
|
144
|
+
```
|
|
145
|
+
|
|
146
|
+
Journal: *Reliability Engineering & System Safety* (2026), article 113057
|
|
147
|
+
Preprint: https://arxiv.org/abs/2604.22755
|
|
148
|
+
|
|
149
|
+
---
|
|
150
|
+
|
|
151
|
+
## License
|
|
152
|
+
|
|
153
|
+
Copyright 2026 Zavier N. Ndum
|
|
154
|
+
|
|
155
|
+
This project is licensed under the Apache License 2.0. See the [LICENSE](https://github.com/SmartLabNuclear/RADIANT_LLM/blob/main/LICENSE) file in the RADIANT_LLM repository for the full license text.
|
|
156
|
+
|
|
157
|
+
|
|
@@ -4,7 +4,7 @@ build-backend = "setuptools.build_meta"
|
|
|
4
4
|
|
|
5
5
|
[project]
|
|
6
6
|
name = "visual-parser"
|
|
7
|
-
version = "2.0.
|
|
7
|
+
version = "2.0.4"
|
|
8
8
|
description = "Standalone Visual-RAG PDF Parser - text extraction and Vision-LLM figure descriptions to JSONL"
|
|
9
9
|
readme = "README.md"
|
|
10
10
|
requires-python = ">=3.10"
|
|
@@ -1,6 +1,6 @@
|
|
|
1
1
|
Metadata-Version: 2.4
|
|
2
2
|
Name: visual-parser
|
|
3
|
-
Version: 2.0.
|
|
3
|
+
Version: 2.0.4
|
|
4
4
|
Summary: Standalone Visual-RAG PDF Parser - text extraction and Vision-LLM figure descriptions to JSONL
|
|
5
5
|
Author: Zavier N. Ndum
|
|
6
6
|
Author-email: zavier.ndum@tamu.edu
|
|
@@ -43,7 +43,6 @@ Requires-Dist: mypy; extra == "dev"
|
|
|
43
43
|
|
|
44
44
|
# visual-parser (Standalone Visual-RAG PDF Ingestion)
|
|
45
45
|
|
|
46
|
-
<!--  -->
|
|
47
46
|

|
|
48
47
|
<!--  -->
|
|
49
48
|
|
|
@@ -75,15 +74,16 @@ Prebuilt images are on **[zev94/radiant-llm](https://hub.docker.com/r/zev94/radi
|
|
|
75
74
|
|
|
76
75
|
| Tag | Description |
|
|
77
76
|
|-----|-------------|
|
|
78
|
-
| `visual-parser-
|
|
79
|
-
| `visual-parser-
|
|
77
|
+
| `visual-parser-latest` | Always latest build (rolling) |
|
|
78
|
+
| `visual-parser-2.0.0` | Pinned release (Apache 2.0 release) |
|
|
79
|
+
| `visual-parser-1.0` | Legacy — v1.0.0, stale |
|
|
80
80
|
|
|
81
81
|
### 1) Install Docker
|
|
82
82
|
- Docker Desktop (Windows/macOS) or Docker Engine (Linux)
|
|
83
83
|
|
|
84
84
|
### 2) Pull the image
|
|
85
85
|
```bash
|
|
86
|
-
docker pull zev94/radiant-llm:visual-parser-
|
|
86
|
+
docker pull zev94/radiant-llm:visual-parser-latest
|
|
87
87
|
```
|
|
88
88
|
|
|
89
89
|
### 3) Run (input + output on the same mounted folder)
|
|
@@ -91,7 +91,7 @@ Windows PowerShell:
|
|
|
91
91
|
```powershell
|
|
92
92
|
docker run --rm --env-file .env `
|
|
93
93
|
-v "C:\path\to\pdfs:/data" `
|
|
94
|
-
zev94/radiant-llm:visual-parser-
|
|
94
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
95
95
|
--input-dir /data --output-dir /data
|
|
96
96
|
```
|
|
97
97
|
|
|
@@ -99,7 +99,7 @@ Linux / WSL:
|
|
|
99
99
|
```bash
|
|
100
100
|
docker run --rm --env-file .env \
|
|
101
101
|
-v "/path/to/pdfs:/data" \
|
|
102
|
-
zev94/radiant-llm:visual-parser-
|
|
102
|
+
zev94/radiant-llm:visual-parser-latest \
|
|
103
103
|
--input-dir /data --output-dir /data
|
|
104
104
|
```
|
|
105
105
|
|
|
@@ -109,7 +109,7 @@ Windows PowerShell:
|
|
|
109
109
|
docker run --rm --env-file .env `
|
|
110
110
|
-v "C:\path\to\pdfs:/data" `
|
|
111
111
|
-v "C:\path\to\out:/out" `
|
|
112
|
-
zev94/radiant-llm:visual-parser-
|
|
112
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
113
113
|
--input-dir /data --output-dir /out
|
|
114
114
|
```
|
|
115
115
|
|
|
@@ -122,11 +122,11 @@ docker images # use the tag printed by Docker
|
|
|
122
122
|
|
|
123
123
|
### Model overrides (optional)
|
|
124
124
|
|
|
125
|
-
Default vision model is **GPT-5.
|
|
125
|
+
Default vision model is **GPT-5.4** when using `--vision-provider gpt`. Override on the command line:
|
|
126
126
|
|
|
127
127
|
```powershell
|
|
128
128
|
docker run --rm --env-file .env -v "C:\path\to\pdfs:/data" `
|
|
129
|
-
zev94/radiant-llm:visual-parser-
|
|
129
|
+
zev94/radiant-llm:visual-parser-latest `
|
|
130
130
|
--input-dir /data --output-dir /data --vision-model gpt-5.4
|
|
131
131
|
```
|
|
132
132
|
|
|
@@ -142,7 +142,7 @@ python visual-parser.py --input-dir "C:\path\to\pdfs"
|
|
|
142
142
|
After pulling the image, run:
|
|
143
143
|
|
|
144
144
|
```bash
|
|
145
|
-
docker run --rm zev94/radiant-llm:visual-parser-
|
|
145
|
+
docker run --rm zev94/radiant-llm:visual-parser-latest --help
|
|
146
146
|
```
|
|
147
147
|
|
|
148
148
|
For copy-paste **Docker** examples (vision presets, text modes, workers, rebuild), see [`docker-usage-examples.md`](docker-usage-examples.md).
|
|
@@ -173,23 +173,28 @@ Performance / misc:
|
|
|
173
173
|
|
|
174
174
|
## Citation
|
|
175
175
|
|
|
176
|
-
If you use RADIANT-LLM or the accompanying evaluation materials, please cite the
|
|
176
|
+
If you use RADIANT-LLM or the accompanying evaluation materials, please cite the journal article:
|
|
177
177
|
|
|
178
178
|
```bibtex
|
|
179
|
-
@article{
|
|
180
|
-
title={
|
|
179
|
+
@article{ndum2026retrieval,
|
|
180
|
+
title={A retrieval-augmented, domain-intelligent agentic framework for reliable decision support in safety-critical nuclear engineering},
|
|
181
181
|
author={Ndum, Zavier Ndum and Tao, Jian and Ford, John and Yim, Mansung and Liu, Yang},
|
|
182
|
-
journal={
|
|
183
|
-
|
|
182
|
+
journal={Reliability Engineering \& System Safety},
|
|
183
|
+
pages={113057},
|
|
184
|
+
year={2026},
|
|
185
|
+
publisher={Elsevier}
|
|
184
186
|
}
|
|
185
187
|
```
|
|
186
188
|
|
|
189
|
+
Journal: *Reliability Engineering & System Safety* (2026), article 113057
|
|
187
190
|
Preprint: https://arxiv.org/abs/2604.22755
|
|
188
191
|
|
|
189
192
|
---
|
|
190
193
|
|
|
191
194
|
## License
|
|
192
195
|
|
|
193
|
-
|
|
196
|
+
Copyright 2026 Zavier N. Ndum
|
|
197
|
+
|
|
198
|
+
This project is licensed under the Apache License 2.0. See the [LICENSE](https://github.com/SmartLabNuclear/RADIANT_LLM/blob/main/LICENSE) file in the RADIANT_LLM repository for the full license text.
|
|
194
199
|
|
|
195
200
|
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|
|
File without changes
|