Skip to main content

Understand - Spatial Intelligence for Real-World AI | Niantic Spatial, Inc.

Hear more about how Niantic Spatial is working with NVIDIA and Flexion to close the Sim-to-Real gap!

Understand (理解)

ピクセルレベルでジオレファレンスされた空間知能。 セグメンテーションと文脈理解を組み合わせることで、人、ロボット、そして大規模言語モデルやAIエージェントにリアルタイムで状況を提供し、より賢く、より効率的な意思決定を可能にします。

オープンセット・セマンティクスに基づき、Niantic Spatial Understand は、3D空間上の各ポイントに意味情報を付与します。これにより、人間と機械の双方が解釈可能な、コンテキスト豊富でクエリ可能な3Dマップを実現します。

Semantic Understanding

次世代セマンティクスの構築

ポイント単位のセマンティック理解 各空間ポイントは、形状、材質、そして文脈的特徴を捉えるオープンかつ高次元の記述子をエンコードします。

オープンセット汎化: 固定的なカテゴリ分類に縛られることなく、再トレーニングなしで未知の概念を認識し、柔軟に適応します。

クロスモーダル・グラウンディング: 視覚・空間・言語の情報を統合し、共通の意味体系のもとで関連付けます。

連続的セマンティクス: オブジェクト、表面、環境にまたがる関係性をきめ細かく理解し、より高度な空間推論を実現します。

Niantic Spatial Understand (理解) の活用

Understand (理解) が可能にすること

クエリ可能なマップ: 「電線や道路で、植生が干渉している場所は?」「車両やドローンが走行・飛行可能な地形はどのようなものか?」「作業現場で遮蔽物や日陰を提供する構造物はどれ?」 といったオープンエンドな空間的質問に対応します。

コンテキスト認識型の自律性: ロボットやAIエージェントが、単なる幾何情報ではなく“意味”に基づいて周囲環境を推論します。

Adaptive intelligence: Maps evolve with new data, improving recognition and inference over time.

Built for integration: Designed for multi-modal queries across enterprise systems, Physical AI agents, and in-field robotics workflows.

Built for the Real World

Defense

Real-time scene interpretation, situational awareness, and mission-critical reasoning for safer, more informed operations.

Robotics & Autonomy

Semantic navigation, task recognition, and adaptive decision-making powered by spatial understanding.

Intelligent Field Operations

Automates inspection and asset intelligence for oil and gas, utilities, and large-scale infrastructure environments.

Frequently Asked Questions

What is semantic understanding in 3D?
+

3D semantic understanding adds meaning and context to a reconstructed environment, not just geometry. Niantic Spatial Understand encodes each 3D point with information about features such as geometry, materials, and context, creating maps that humans, machines, and AI agents can query, measure, and reason over.

What are open-set semantics?
+

Open-set semantics allow a system to identify and reason about concepts beyond a fixed list of predefined categories. In Understand, open-set generalization is designed to adapt to new concepts without retraining for every new category or relying only on a closed taxonomy. That makes the semantic layer better suited to open-ended questions about real environments, including objects, materials, relationships, and conditions.

What is per-point semantic understanding?
+

Per-point semantic understanding attaches a semantic descriptor to each point in a 3D scene. The descriptor can capture geometric, material, and contextual features, allowing the system to search or compare specific parts of an environment rather than treating the entire scene as one label. This creates a spatially grounded foundation for segmentation, measurement, natural-language search, and downstream AI workflows.

Can you query a 3D map using natural language?
+

Yes. It is designed to support open-ended, multimodal questions about a 3D environment, such as where vegetation encroaches on infrastructure, which terrain is traversable, or which structures provide shelter or shade. Cross-modal grounding connects visual, spatial, and linguistic information so the answer can be tied to a location or region in the mapped environment.

How is 3D semantic understanding different from traditional object detection?
+

Traditional object detection or closed-set segmentation generally assigns predefined labels to visible regions. 3D semantic understanding adds spatial grounding and context: it can represent meaning at individual 3D points, connect relationships across objects and surfaces, and support open-ended queries. Understand is positioned as a layer for interpreting and operating in a measured 3D environment, not simply labeling isolated 2D images.

What can 3D semantic understanding be used for?
+

3D semantic understanding can support scene interpretation, situational awareness, semantic navigation, task recognition, inspection, and asset intelligence. Humans, robots, and AI agents in can use meaning, not geometry alone, to interpret surroundings, identify relevant areas, reason about conditions, and make more informed decisions.

How does spatial semantics help AI agents and robots?
+

Spatial semantics gives AI agents and robots a shared, queryable representation of the physical world. By grounding visual, spatial, and linguistic information in a 3D map, it helps machines interpret their surroundings and act on context, not just geometry. That foundation supports applications such as semantic navigation, task recognition, adaptive decision-making, inspection, and asset intelligence.