A geometric vocabulary extractor that reads structural properties from latent patches — and proved that text carries the same geometric structure as images.
This is a two-tier gated geometric transformer trained on 27 geometric primitives (point through channel) in 8×16×16 voxel grids. It extracts 17-dimensional gate vectors (explicit geometric properties) and 256-dimensional patch features (learned representations) from any compatible latent input.
What It Does
Takes an
(8, 16, 16)
tensor — originally voxel grids, but proven to work on adapted FLUX VAE latents and text-derived latent patches — and produces per-patch geometric descriptors:
Dimensions 0–10 are
local
(intrinsic to each patch, no cross-patch info). Dimensions 11–16 are
structural
(relational, computed after attention sees neighborhood context).
Architecture
(8, 16, 16) input
↓
PatchEmbedding3D → (B, 64, 64) # 64 patches of 32 voxels each
↓
Stage 0: Local Encoder + Gate Heads # dims, curvature, boundary, axes
↓
proj([embedding, local_gates]) → (B, 64, 128)
↓
Stage 1: Bootstrap Transformer ×2 # standard attention with local context
↓
Stage 1.5: Structural Gate Heads # topology, neighbors, surface role
↓
Stage 2: Geometric Transformer ×2 # gated attention modulated by all 17 gates
↓
Stage 3: Classification Heads # 27-class shape recognition
The geometric transformer blocks use gate-modulated attention: Q and K are projected from
[hidden, all_gates]
, V is multiplicatively gated, and per-head compatibility scores are computed from gate interactions.
The Rosetta Stone Discovery
This model was used as the analyzer in the
GeoVAE Proto experiments
, which proved that text descriptions produce
2.5–3.5× stronger geometric differentiation
than actual images when projected through a lightweight VAE into this model's patch space.
Source
patch_feat discriminability
FLUX images (49k)
+0.020
flan-t5-small text
+0.053
bert-base-uncased text
+0.053
bert-beatrix-2048 text
+0.050
Three architecturally different text encoders converge to ±5% of each other — the geometric structure is in the language, not the encoder. This model reads it.
Training
Trained on procedurally generated multi-shape superposition grids (2–4 overlapping geometric primitives per sample, 27 shape classes). Two-tier gate supervision with ground truth computed from voxel analysis:
Local gates
: dimensionality from axis extent, curvature from fill ratio, boundary from partial occupancy
Structural gates
: topology from 3D convolution neighbor counting, surface role from neighbor density thresholds
200 epochs, achieving 93.8% recall on shape classification with explicit geometric property prediction as auxiliary objectives.
Files
File
Description
geometric_model.py
Standalone model +
load_from_hub()
+
extract_features()
model.pt
Pretrained weights (epoch 200)
Usage
import torch
from geometric_model import SuperpositionPatchClassifier, load_from_hub, extract_features
# Load pretrained
model = load_from_hub()
# From any (8, 16, 16) source
patches = torch.randn(16, 8, 16, 16).cuda()
gate_vectors, patch_features = extract_features(model, patches)
# Or full output dict
out = model(patches)
out["local_dim_logits"] # (B, 64, 4) dimensionality
out["local_curv_logits"] # (B, 64, 3) curvature
out["struct_topo_logits"] # (B, 64, 2) topology
out["patch_features"] # (B, 64, 128) learned features
out["patch_shape_logits"] # (B, 64, 27) shape classification
Geometric deep learning research by
AbstractPhil
. The model demonstrates that geometric structure is a universal language bridging text and visual modalities — symbolic association through geometric language.
Runs of AbstractPhil geovocab-patch-maker on huggingface.co
4
Total runs
0
24-hour runs
0
3-day runs
2
7-day runs
4
30-day runs
More Information About geovocab-patch-maker huggingface.co Model
geovocab-patch-maker huggingface.co is an AI model on huggingface.co that provides geovocab-patch-maker's model effect (), which can be used instantly with this AbstractPhil geovocab-patch-maker model. huggingface.co supports a free trial of the geovocab-patch-maker model, and also provides paid use of the geovocab-patch-maker. Support call geovocab-patch-maker model through api, including Node.js, Python, http.
geovocab-patch-maker huggingface.co is an online trial and call api platform, which integrates geovocab-patch-maker's modeling effects, including api services, and provides a free online trial of geovocab-patch-maker, you can try geovocab-patch-maker online for free by clicking the link below.
AbstractPhil geovocab-patch-maker online free url in huggingface.co:
geovocab-patch-maker is an open source model from GitHub that offers a free installation service, and any user can find geovocab-patch-maker on GitHub to install. At the same time, huggingface.co provides the effect of geovocab-patch-maker install, users can directly use geovocab-patch-maker installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
geovocab-patch-maker install url in huggingface.co: