@thinkingos/vsl-sdk 0.1.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
package/LICENSE.txt ADDED
@@ -0,0 +1,12 @@
1
+ Creative Commons Attribution 4.0 International (CC BY 4.0)
2
+
3
+ You are free to:
4
+ - Share — copy and redistribute the material in any medium or format
5
+ - Adapt — remix, transform, and build upon the material for any purpose, even commercially.
6
+
7
+ Under the following terms:
8
+ - Attribution — You must give appropriate credit, provide a link to the license, and indicate if changes were made.
9
+
10
+ No additional restrictions — You may not apply legal terms or technological measures that legally restrict others from doing anything the license permits.
11
+
12
+ Full license text: https://creativecommons.org/licenses/by/4.0/
package/README.md ADDED
@@ -0,0 +1,84 @@
1
+ 🧾 Defensive Publication: Visual Scene Language (VSL)
2
+
3
+ Author: Maxim Zhadobin
4
+ Publication Date: 28.05.2025
5
+ Version: v0.1
6
+
7
+ ⸻
8
+
9
+ 📘 Invention Title
10
+
11
+ Method for Representing and Editing Visual Scenes for Language Models via a Structured Format: Visual Scene Language (VSL)
12
+
13
+ ⸻
14
+
15
+ 📌 Application Domains
16
+ • Artificial Intelligence and Large Language Models (LLMs)
17
+ • Image generation and editing
18
+ • UX/UI design
19
+ • Architecture, Engineering, and Construction (AEC)
20
+ • 2D/3D modeling, AR/VR
21
+
22
+ ⸻
23
+
24
+ 🧠 Abstract
25
+
26
+ A method is proposed for describing visual scenes in a machine-readable format interpretable by language models (LLMs). The format is a structured JSON schema containing:
27
+ • canvas parameters (dimensions, measurement units, background),
28
+ • a list of objects (type, size, coordinates, anchor points),
29
+ • optional styles and object relationships.
30
+
31
+ An LLM can use this structure:
32
+ • to understand the image as a meaningful scene,
33
+ • to edit the scene based on textual commands,
34
+ • to generate a new scene,
35
+ • to reconstruct an image from the JSON.
36
+
37
+ ⸻
38
+
39
+ 🔄 Technological Workflow
40
+ 1. An image is converted into a JSON-based scene structure (VSL).
41
+ 2. The LLM interprets and modifies the JSON in response to a text prompt.
42
+ 3. The modified JSON is rendered into a new image.
43
+ 4. The cycle may repeat iteratively.
44
+
45
+ ⸻
46
+
47
+ 🧩 Distinctive Features
48
+ • Unlike generative models (DALL·E, Midjourney), this method separates the semantic structure of the scene from its visual rendering.
49
+ • Unlike scene graphs in computer graphics, VSL is optimized for language models and semantic processing.
50
+ • The method does not require a visual interface — interaction occurs through structure and natural language commands.
51
+
52
+ ⸻
53
+
54
+ 💡 Example Applications
55
+ • Editing user interfaces, slides, and illustrations without a visual editor.
56
+ • Generating AR/VR scenes from textual scenarios.
57
+ • Interactive assistants working with visual objects.
58
+ • Robotics: LLM perceives the environment through VSL and gives structured commands.
59
+ • Architectural and engineering design: users describe a building or structure via text; the LLM generates a corresponding VSL structure. A renderer builds a 2D/3D model from it. Edits are made via natural language and reflected in the scene. The system can also validate design choices against regulatory standards.
60
+
61
+ ⸻
62
+
63
+ 📜 Legal Status
64
+
65
+ This document is published with the intent to prevent patent claims by third parties. The author waives exclusive patent rights in favor of public use but formally asserts authorship, concept origin, and publication date.
66
+
67
+ ⸻
68
+
69
+ 📎 Attachments (to be included in GitHub repository)
70
+ • vsl_scene_example.json: example scene with a red rectangle
71
+ • llm_prompts_examples.md: prompt examples for modifying the scene
72
+ • vsl_to_image.py: Python script to visualize the JSON scene
73
+ • vsl_editor_mockup.ipynb: (optional) Jupyter Notebook mock editor
74
+ • Extended format specs: styles, materials, animation, 3D coordinates, behavioral logic
75
+
76
+ ⸻
77
+
78
+ 🔖 License
79
+
80
+ Creative Commons Attribution 4.0 International (CC BY 4.0)
81
+
82
+ ⸻
83
+
84
+ “Visual Scene Language is not just a way to describe an image — it is a language for spatial thinking by artificial intelligence.”