GLACIER AI · AI TOOLS

Find the right AI Tools.

From video creation to text processing, explore 84 capabilities across 6 AI categories.
Search by task to find the right tool.

Explore AI tools

All · 84 capabilities, showing 12

Need an integration? Contact us ↗
VIDEO

Text-to-video and image-to-video

Generate video clips from text prompts or reference images for creative content, ad assets and previsualization.

VIDEO

Video super-resolution and frame interpolation

Upscale video resolution and frame rate for smoother, sharper footage, ideal for restoring old films and slow motion.

VIDEO

Video highlight detection and auto-editing

Automatically find the best moments and cut them into short clips for sports, live streams and short-form video.

VIDEO

Video subtitle removal

Detect and remove burned-in subtitles for re-dubbing and multilingual distribution.

VIDEO

Video watermark removal

Locate and remove watermarks and channel logos for asset cleanup and compliance.

VIDEO

Video stabilization

Remove handheld camera shake for steady footage and a better viewing experience.

VIDEO

Video matting and background replacement

Separate people in the foreground from the background, enabling virtual backgrounds without a green screen.

VIDEO

Action and event recognition

Recognize human actions and events in video for security, sports and industrial monitoring.

VIDEO

Activity, fall and fatigue detection

Detect falls, fatigue and abnormal activity in elder care, construction sites and driving, and raise alerts.

VIDEO

Visual anomaly detection

Spot deviations from normal patterns in video for production lines, inspections and security.

VIDEO

Multi-object tracking

Continuously track multiple objects in video and output track IDs for pedestrian and traffic flow analysis.

VIDEO

Object detection and multi-object tracking

Detect objects in video and images and track them over time, powering counting, speed measurement and behavior analysis.

VIDEO

Trajectory prediction and motion planning

Predict the future paths of pedestrians, vehicles and robots for autonomous driving and robot obstacle avoidance.

VIDEO

Lip sync and digital human presenters

Align a digital human's lip movements with speech for virtual hosts, newscasts and customer service.

VIDEO

Vision-language-action and grasping

Drive robotic grasping from vision and language instructions for embodied AI.

TEXT

General writing and chat

Text generation, rewriting, Q&A and multi-turn chat for general assistant use cases.

TEXT

Summarization and rewriting

Condense long text, rewrite in a new style and extract key points for reports, news and meeting notes.

TEXT

Machine translation and localization

Translate between languages and polish for local audiences across documents, interfaces and subtitles.

TEXT

Text content moderation

Flag violating, violent or extremist, sexual and spam text for communities and content platforms.

TEXT

Text embeddings and semantic similarity

Encode text as vectors for semantic search, deduplication, clustering and similarity matching.

TEXT

Entity and relation extraction

Extract people, organizations, events and their relationships from text to build knowledge graphs.

TEXT

Sentiment and opinion analysis

Determine sentiment polarity and opinion in text for public opinion monitoring, reviews and customer service QA.

TEXT

Intent recognition and slot filling

Understand user intent and extract key slots for dialogue systems and ticket routing.

TEXT

Tabular classification and risk scoring

Classify tables and forms and output risk scores for risk control, compliance and review.

TEXT

Field and form extraction

Extract structured fields from contracts, receipts and forms to automate data entry.

TEXT

Text language identification

Automatically identify the language of text for routing, pre-translation and content distribution.

TEXT

Language identification

Identify the language of speech or text for multilingual customer service and meeting systems.

TEXT

PII/PHI detection and redaction

Detect and redact personal and health information to meet privacy compliance.

TEXT

AI security detection and response

Identify risks such as prompt injection, jailbreaks and harmful content, and trigger response policies.

TEXT

Causal inference and uplift modeling

Estimate causal effects and incremental gains of interventions for marketing, operations and pricing.

TEXT

Graphs, anti-fraud and relationship risk

Mine fraud rings and linked risks from graphs for anti-fraud and risk control.

IMAGE

Text-to-image and reference-based generation

Generate images from text or reference images for creative work, design and marketing assets.

IMAGE

Inpainting, outpainting and controllable editing

Repaint regions, extend beyond the frame and edit with precise control for retouching, e-commerce and design.

IMAGE

Image classification and fine-grained recognition

Recognize image categories and fine-grained subclasses for quality inspection, archiving and identification.

IMAGE

OCR text recognition

Recognize text in images, including receipts, ID documents, documents and scene text.

IMAGE

Layout, table and formula parsing

Parse document layouts, table structures and math formulas for document digitization.

IMAGE

Face detection, recognition and liveness

Face detection, matching and liveness checks for access control, identity verification and security.

IMAGE

Pose, gesture and keypoints

Recognize body pose, hand gestures and keypoints for interaction, fitness and animation.

IMAGE

Medical image detection and segmentation

Detect and segment lesions and organs for diagnostic support and research.

IMAGE

Supervised defect detection

Detect surface and internal product defects from labeled samples for industrial quality inspection.

IMAGE

Visual measurement and alignment

Measure dimensions, position and alignment for precision manufacturing and assembly.

IMAGE

Open-vocabulary detection & segmentation

Detect and segment objects described in natural language, with no fixed classes.

IMAGE

Image & video segmentation

Semantic, instance and panoptic segmentation for matting, labeling and analysis.

IMAGE

3D reconstruction & digital twins

Reconstruct 3D scenes from images or point clouds for digital twins, mapping and simulation.

IMAGE

Depth estimation

Monocular and stereo depth estimation for autonomous driving, AR and robotics.

IMAGE

Super-resolution, denoising & deblurring

Image enhancement and restoration for sharper quality and old photo repair.

IMAGE

Text/image to 3D assets

Generate 3D models and assets from text or images for games, e-commerce and the metaverse.

SPEECH / AUDIO

Speech-to-text ASR

Transcribe speech into text for meetings, customer service and subtitles.

SPEECH / AUDIO

Text-to-speech TTS

Turn text into natural speech for announcements, audio content and accessibility.

SPEECH / AUDIO

Real-time voice agents

Real-time voice conversations that take business actions, for call centers and voice assistants.

SPEECH / AUDIO

Speech-to-speech translation

End-to-end speech translation that keeps voice and prosody, for cross-language communication.

SPEECH / AUDIO

Voice cloning & conversion

Clone a voice or convert speaking style for dubbing and personalized TTS.

SPEECH / AUDIO

Speaker diarization

Tell speakers apart for meeting notes and multi-speaker transcription.

SPEECH / AUDIO

Voiceprint recognition & verification

Confirm identity by voiceprint for identity checks and fraud prevention.

SPEECH / AUDIO

Wake word detection

Detect a chosen wake word for low-power voice assistant activation.

SPEECH / AUDIO

Voice activity detection VAD

Separate speech from non-speech to cut costs, segment audio and power real-time interaction.

SPEECH / AUDIO

Noise reduction & speech enhancement

Suppress noise and enhance voices for calls, recordings and meetings.

SPEECH / AUDIO

Echo cancellation & dereverberation

Remove echo and reverb for hands-free calls and meeting room audio.

SPEECH / AUDIO

Environmental sound recognition

Recognize sound events around you for security, industry and urban sensing.

SPEECH / AUDIO

Music & sound effect generation

Generate music and sound effects for content creation, games and short videos.

MULTIMODAL

RAG knowledge Q&A

Retrieval-augmented generation that answers from your enterprise knowledge base with fewer hallucinations.

MULTIMODAL

Document Q&A & comparison

Ask, compare and summarize across documents such as contracts, reports and papers.

MULTIMODAL

Enterprise knowledge assistant

An all-in-one assistant for internal knowledge search, Q&A and writing.

MULTIMODAL

Multimodal retrieval

Cross-modal search: find images or videos by text, or text by image.

MULTIMODAL

Visual Q&A & scene understanding

Answer questions about images for accessibility, education and inspections.

MULTIMODAL

Real-time multimodal assistant

An interactive assistant that handles speech, images and text in real time.

MULTIMODAL

Tool-calling & workflow agents

Call external tools and APIs to orchestrate workflows and complete complex tasks.

MULTIMODAL

GUI & browser agents

Operate graphical interfaces and browsers to finish tasks, for automation and RPA.

MULTIMODAL

Coding & DevOps agents

Code generation, understanding and review, plus operations automation.

MULTIMODAL

Code generation & understanding

Code completion, generation, explanation, refactoring and unit test generation.

MULTIMODAL

Multi-agent orchestration for research & ops

Multiple agents working together on research, operations and analysis tasks.

MULTIMODAL

Personalized ranking

Rank results by user behavior for recommendations, search and ads.

MULTIMODAL

Candidate retrieval & similar items

Retrieve candidates and recommend similar items for e-commerce, content and social.

MULTIMODAL

Semantic search & reranking

Semantic retrieval + reranking for more relevant search results.

TIME SERIES / SENSING

Predictive maintenance & remaining useful life

Predict equipment failures and remaining useful life from multi-source data.

TIME SERIES / SENSING

Streaming anomaly & change detection

Detect anomalies and sudden changes in streaming data in real time.

TIME SERIES / SENSING

Time series forecasting

Forecast trends, seasonality and anomalies in time series data.

TIME SERIES / SENSING

Scheduling, routing & resource optimization

Solve scheduling, routing and resource allocation optimization problems.

TIME SERIES / SENSING

AMR navigation & dispatch

Mobile robot navigation, obstacle avoidance and fleet dispatch.

TIME SERIES / SENSING

Drone & on-site autonomous inspection

Autonomous drone inspection, recognition and alerts.

TIME SERIES / SENSING

Multi-sensor perception & fusion

Fuse cameras, radar, LiDAR and other sensors for perception.

TIME SERIES / SENSING

IMU/radar/LiDAR fusion

Fuse inertial, radar and LiDAR data for localization and perception.

TIME SERIES / SENSING

Visual-inertial SLAM

Visual + inertial odometry and mapping for robotics and AR.

TIME SERIES / SENSING

Load & environmental forecasting

Forecast power load and environmental metrics for energy and city management.

FEATURED / Quality enhancement

More detail in every frame.

Super-resolution, quality enhancement, AI denoising and bitrate recovery. See real-world results on original videos and images.

Try the quality demos ↗
3 demo types3 video comparisons17 image pairs
Frame from a super-resolution video▷ Watch the original video demo
Start from the task

Make AI part of your team.

Find a capability that interests you, then talk to us about use cases and integration.

GLACIER AI

Start with what you need.

Tell us which products interest you, your current systems and your goals.

shenzhongchang@glacier.fit

Product scope, configuration and activation are confirmed with you directly.