Trust score
0 / 100
Track record
0 / 100
Thesis calls
76
Evaluated calls
0
Average return
n/a
Win rate
n/a

Past bets that played out

These are the clearest thesis calls with observable outcomes, linked back to the original videos.

PLTRopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 38 / 100
Source: Can LLMs Introspect? A Reality Check
SNOWopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 40 / 100
Source: Can LLMs Introspect? A Reality Check
GOOGLopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 50 / 100
Source: Can LLMs Introspect? A Reality Check

Latest videos and market context

Recent source posts from this author. Create an account to inspect the complete persisted research trail.

Can LLMs Introspect? A Reality Check

May 27, 2026, 12:00 AM EDT

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization

May 27, 2026, 12:00 AM EDT

Academic paper proposes a geometry-conditioned autoregressive model to generate *physically buildable* brick assemblies (stability + discrete parts) from 3D inputs using point clouds, structure-aware tokenization, and constrained decoding/rollback. If commercialized, it primarily strengthens the “AI-assisted 3D/CAD/content creation” toolchain and simulation-driven design workflows; direct public-market impact is most plausible via GPU/AI infrastructure and 3D/CAD software platforms rather than toy manufacturers (LEGO is private).

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

Jun 3, 2026, 12:00 AM EDT

AURA-Mem proposes action-gated, constant-size recurrent memory for long-horizon embodied/robot policies on bandwidth- and memory-constrained edge hardware. If it (or similar methods) becomes standard in robotics VLA stacks, it shifts the bottleneck from “more VRAM / more memory bandwidth” toward “smarter memory-write policies,” potentially enabling cheaper edge deployments and improving flash endurance. Near-term investability is indirect: it’s a research result (early arXiv) without announced product adoption, but it is directionally relevant to edge AI/robotics compute, memory/flash endurance, and robotics platform economics.

Visual Graph Scaffolds for Structural Reasoning in Large Language Models

Jun 3, 2026, 12:00 AM EDT

Paper claims visual graph-structured “mind map” scaffolds materially improve LLM multi-hop reasoning under “abstract guidance” (no direct answer hints), outperforming flattened text graph representations; benefits persist post SFT and KL distillation. Investable implication is incremental tailwind for multimodal/vision-language model stacks and tooling that enable structured visual reasoning and UI-level reasoning scaffolds, but it is early-stage and not yet a clear product catalyst on its own.

Proof-backed call history

These are recent thesis calls tied to original source content where available.

PLTRopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 38 / 100
Source: Can LLMs Introspect? A Reality Check
SNOWopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 40 / 100
Source: Can LLMs Introspect? A Reality Check
GOOGLopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 50 / 100
Source: Can LLMs Introspect? A Reality Check
MSFTopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 53 / 100
Source: Can LLMs Introspect? A Reality Check
CRWDopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 54 / 100
Source: Can LLMs Introspect? A Reality Check
PANWopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 56 / 100
Source: Can LLMs Introspect? A Reality Check
DDOGopen

Paper argues prior “LLM introspection” results are likely confounded by surface-cue pattern matching; behavioral tests alone don’t prove privileged access to internal states. Better-controlled relabeling drops performance toward chance. Market implication: de-risks hype around near-term ‘self-diagnosing’/self-auditing models; increases need for external monitoring, eval, governance, and tooling rather than relying on model self-reports.

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 58 / 100
Source: Can LLMs Introspect? A Reality Check
SNOWopen

Academic paper proposes a geometry-conditioned autoregressive model to generate *physically buildable* brick assemblies (stability + discrete parts) from 3D inputs using point clouds, structure-aware tokenization, and constrained decoding/rollback. If commercialized, it primarily strengthens the “AI-assisted 3D/CAD/content creation” toolchain and simulation-driven design workflows; direct public-market impact is most plausible via GPU/AI infrastructure and 3D/CAD software platforms rather than t

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 22 / 100
Source: BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization
ADBEopen

Academic paper proposes a geometry-conditioned autoregressive model to generate *physically buildable* brick assemblies (stability + discrete parts) from 3D inputs using point clouds, structure-aware tokenization, and constrained decoding/rollback. If commercialized, it primarily strengthens the “AI-assisted 3D/CAD/content creation” toolchain and simulation-driven design workflows; direct public-market impact is most plausible via GPU/AI infrastructure and 3D/CAD software platforms rather than t

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 30 / 100
Source: BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization
PTCopen

Academic paper proposes a geometry-conditioned autoregressive model to generate *physically buildable* brick assemblies (stability + discrete parts) from 3D inputs using point clouds, structure-aware tokenization, and constrained decoding/rollback. If commercialized, it primarily strengthens the “AI-assisted 3D/CAD/content creation” toolchain and simulation-driven design workflows; direct public-market impact is most plausible via GPU/AI infrastructure and 3D/CAD software platforms rather than t

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 34 / 100
Source: BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization
RBLXopen

Academic paper proposes a geometry-conditioned autoregressive model to generate *physically buildable* brick assemblies (stability + discrete parts) from 3D inputs using point clouds, structure-aware tokenization, and constrained decoding/rollback. If commercialized, it primarily strengthens the “AI-assisted 3D/CAD/content creation” toolchain and simulation-driven design workflows; direct public-market impact is most plausible via GPU/AI infrastructure and 3D/CAD software platforms rather than t

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 33 / 100
Source: BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization
Uopen

Academic paper proposes a geometry-conditioned autoregressive model to generate *physically buildable* brick assemblies (stability + discrete parts) from 3D inputs using point clouds, structure-aware tokenization, and constrained decoding/rollback. If commercialized, it primarily strengthens the “AI-assisted 3D/CAD/content creation” toolchain and simulation-driven design workflows; direct public-market impact is most plausible via GPU/AI infrastructure and 3D/CAD software platforms rather than t

Mentioned: May 27, 2026, 12:00 AM EDTConviction: 38 / 100
Source: BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization

About this channel

Channel bio, source link, and public-market context from YouTube.

Subscribersn/a
Videosn/a
Win raten/a
Average returnn/a

arXiv cs.AI

Unlock the full track record

Create an account to inspect the complete author history, trust-weighted rankings, and persisted research across authors, theses, and assets.

64 more thesis calls are available after sign-up.