BorySense Platform | Bory AI Video Understanding and Conversational Accessibility Platform

Bory's shared AI accessibility platform that understands video, audio, lip movements, and conversation data with AI and converts them into captions, summaries, plain-language explanations, and accessibility APIs

 

BorySense Platform is Bory's AI video understanding and conversational accessibility platform that analyzes video, audio, lip movements, and conversation data with AI and connects them to real-time captions, video captions, conversation summaries, plain-language conversion, consultation records, accessible screens, and external service APIs.

 

Public institutions, hospitals, educational institutions, welfare organizations, corporate customer centers, kiosks, AR glasses, web services, chatbots, and consultation systems rely heavily on information delivered through audio, video, and conversation. However, not every user can hear speech clearly, understand long sentences easily, or immediately follow complex guidance procedures.

 

BorySense Platform uses AI to understand this information and converts it into formats that are easier for users to access. It can turn speech into captions, long conversations into summaries, complex guidance into plain language, and situations in video into accessibility information for use across a wide range of products and services.

 

BorySense Platform can be deployed as a standalone platform or integrated as a shared AI engine across Bory products, solutions, and services.

 

[Platform Implementation Inquiry][AI Accessibility Engine Integration Consultation][PoC Development Inquiry][Request a Proposal]

Platform Definition

 

BorySense Platform is a shared platform that modularizes Bory's AI accessibility technologies.

A product is a packaged offering delivered to a customer, a solution is a system built for the customer's environment, and a platform is a shared technology foundation for repeatedly creating multiple products and solutions.

BorySense Platform provides the following capabilities within a unified platform architecture.

  • Video Understanding AI
  • Speech Recognition STT
  • Lip-Reading STT
  • Real-Time Caption Generation
  • Video Content Captioning
  • Summarize conversations
  • Convert text into plain language
  • Consultation Record Generation
  • Configure accessible UI
  • Integration with Kiosks, AR Glasses, Web, and Apps
  • External Service API Provision
  • Support for On-Premises and Cloud Deployment

In other words, BorySense Platform is Bory's shared AI accessibility platform that “sees, listens, understands, and converts information into easier-to-use formats.”

 

Customer Challenges

 

Audio- and video-centered information is not equally accessible to every user

In public institutions, hospitals, educational institutions, companies, stores, and customer centers, a large amount of information is delivered through audio, video, and conversation.

However, people who are deaf or hard of hearing, older adults, foreign-language users, people with developmental disabilities, and digitally vulnerable users may have difficulty understanding audio-centered information or complex sentences.

BorySense Platform analyzes audio, video, and conversation information with AI and converts it into captions, summaries, and plain-language explanations to improve information accessibility.

 

Speech recognition alone cannot reliably process on-site conversations

At public service counters, hospital reception desks, stores, kiosks, meeting rooms, and classrooms, background noise, speaker distance, rapid speech, pronunciation differences, and simultaneous speech by multiple people may occur.

BorySense Platform can be configured to use lip-movement video and conversational context together with audio information to improve captioning and conversation-understanding quality.

 

Complex guidance text makes information harder for users to understand

Public-service guidance, hospital instructions, financial information, product descriptions, application procedures, terms of use, and test instructions often contain technical terms and long sentences.

BorySense Platform converts complex sentences into plain language and summarizes key points so users can understand the information more easily.

 

Recording and summarizing consultations and meetings creates a significant workload

Customer consultations, public-service consultations, hospital reception, meetings, lectures, and educational consultations require records and organization after the conversation.

BorySense Platform transcribes conversations, summarizes key points, and, where necessary, organizes them into follow-up actions or guidance messages.

 

Developing accessibility features separately for every product is inefficient

If kiosks, AR glasses, chatbots, websites, consultation systems, video platforms, and hospital reception systems are operated independently, captioning, summarization, and plain-language explanation functions must be developed repeatedly.

BorySense Platform provides a shared AI accessibility engine and APIs that can be used across multiple Bory products and solutions, reducing duplicate development and improving scalability.

 

Platform Overview

 

BorySense Platform is an AI accessibility platform that processes video, audio, lip-movement, and conversation data.

The main components are as follows.

  • Video Input Module
  • Audio Input Module
  • Lip-Reading Analysis Module
  • Speech Recognition STT Module
  • Caption Generation Module
  • Conversation summarization module
  • Plain-language conversion module
  • Video Understanding Module
  • Consultation Record Generation Module
  • Accessibility UI Module
  • Administrator Dashboard
  • API integration module
  • Data Security and Access-Control Module

BorySense Platform can be deployed as a standalone platform or integrated as a shared engine across Bory products, solutions, and services.

 

Core Modules

 

1. Video Understanding AI Module

A module that analyzes people, facial regions, lip movements, gaze direction, behavior, scenes, and key objects in video.

It can serve as a foundation for analyzing the information users are viewing and the conversational context at consultation counters, meeting rooms, classrooms, kiosks, AR glasses, and video content.

Examples of use include the following.

  • Lip Region Extraction
  • Speaker Face Region Analysis
  • User Approach Detection
  • Understanding Key Scenes in Video
  • Consultation and Educational Video Analysis
  • Assistance with Accessible Content Generation

 

2. Speech Recognition STT Module

A module that converts speech captured by a microphone or contained in a video file into text.

It can be used for real-time conversations, consultations, meetings, lectures, kiosk orders, customer-center consultations, and video caption generation.

Examples of use include the following.

  • Real-Time Consultation Caption Generation
  • Meeting and Lecture Caption Generation
  • Customer Center Consultation Records
  • Transcription of Kiosk Voice Orders
  • Generate Captions from Video Files
  • Conversational AI Input Data Generation

 

3. Lip-Reading STT Module

A module that uses lip-movement video to transcribe speech or supplement speech-recognition results.

It can provide supplementary information in noisy environments, when speech is faint, or when distance reduces speech-recognition quality.

Examples of use include the following.

  • Hearing Assistance AR Captions
  • Lip-Reading STT Assistance at Public Service Counters
  • Caption Assistance for Meeting and Consultation Environments
  • Video-Based Speech Analysis
  • Audio-and-Lip Combined STT

 

4. Real-Time Caption Generation Module

A module that generates easy-to-read captions based on STT results and conversational context.

Beyond simple text conversion, it can be combined with accessibility UI features such as display length, sentence segmentation, screen readability, caption speed, high-contrast display, and large-text display.

Examples of use include the following.

  • Public-Service Consultation Captions
  • Hospital Reception Captions
  • AR Glasses Captions
  • Kiosk Guidance Captions
  • Real-Time Meeting and Lecture Captions
  • Video Content Captions

 

5. Conversation Summarization Module

A module that organizes long conversations or consultations around their key points.

Users can quickly review major questions, answers, decisions, instructions, and follow-up actions without rereading the entire conversation.

Examples of use include the following.

  • Public-Service Consultation Summary
  • Hospital Reception and Consultation Summary
  • Customer-Center Consultation Summary
  • Meeting Minutes Summary
  • Educational Consultation Record Summary
  • Chatbot Consultation History Summary

 

6. Plain-Language Conversion Module

A module that converts complex guidance, technical terminology, and long sentences into easier expressions.

It helps users understand content more easily in public institutions, hospitals, welfare organizations, kiosks, and educational services.

Examples of use include the following.

  • Plain-Language Explanation of Public-Service Application Procedures
  • Plain-Language Explanation of Hospital Test Instructions
  • Plain-Language Explanation of Kiosk Ordering and Payment
  • Plain-Language Conversion of Product Descriptions
  • Guidance Text Generation for Older Adults and Digitally Vulnerable Users
  • Accessibility Guidance Assistance for People with Developmental Disabilities

 

7. Accessibility UI Module

A module that displays captions, summaries, and plain-language explanations in user-friendly screen formats.

The UI can be configured for large text, high contrast, simplified views, step-by-step guidance, caption-position adjustment, AR glasses displays, kiosk screens, mobile screens, and web screens.

Examples of use include the following.

  • Caption Screen for Visitors
  • AR Glasses Caption UI
  • Simplified Kiosk Screen
  • Hospital Reception Guidance Screen
  • Educational Video Caption Screen
  • Record Review Screen for Consultation Staff

 

8. Accessibility API Module

An API module for integrating BorySense Platform captions, summaries, plain-language text, and video-understanding results with external services.

Customers can maintain their existing service architecture while integrating the accessibility capabilities they need through APIs.

Examples of use include the following.

  • Kiosk Integration API
  • AR Glasses Caption API
  • Chatbot Consultation API
  • Video Platform Caption API
  • Hospital Reception System Integration API
  • Public Institution Civil-Service System Integration API

 

Platform Deployment Models

 

Cloud-Based

A model in which AI analysis engines and APIs are provided through Bory servers or a cloud environment.

Shared accessibility functions can be used across multiple locations, customers, and services.

 

On-Premises Model

In environments with restricted external network access, such as public institutions, hospitals, or corporate internal networks, the platform can be deployed on internal servers or within a closed network.

This model is suitable when sensitive data such as consultation audio, video, civil-service information, or hospital information cannot easily be sent outside the organization.

 

Edge-Integrated

A model that integrates with kiosks, AR glasses, camera equipment, on-site PCs, and board-based devices to process part of the analysis on site.

It can be used where real-time performance, security, and network restrictions are important.

 

API Integration Model

A model that integrates BorySense capabilities through APIs with the customer's existing apps, web services, kiosks, consultation systems, and video platforms.

It allows captions, summaries, plain-language conversion, and video-understanding features to be added while retaining the existing system.

 

Application Areas

 

Accessibility for Public-Service Consultations

At public service counters, staff speech can be displayed as captions and complex administrative guidance can be converted into plain language.

When integrated with BoryGlass Gov Solution, it can create a barrier-free public-service consultation environment for people who are deaf or hard of hearing and for older adults.

 

Hospital Reception and Consultation Accessibility

Hospital reception, payment, test instructions, administrative consultations, and care guidance can be provided as captions and plain-language explanations.

It helps older patients and patients with hearing loss understand hospital procedures more easily.

 

Hearing Assistance with AR Glasses

Real-time conversation captions, meeting captions, lecture captions, and guidance messages can be displayed on AR glasses.

Users can view conversation captions while looking at the other person.

 

Barrier-Free Kiosk Guidance

Kiosks can provide voice ordering, captioned guidance, plain-language guidance, and large-text, high-contrast screens.

It improves kiosk usability for older adults, people with disabilities, foreign-language users, and digitally vulnerable users.

 

Meeting and Lecture Captioning and Summarization

In meetings, lectures, seminars, and educational settings, speech can be displayed as captions and key content can be summarized after the session.

This strengthens both hearing accessibility and record-management capabilities.

 

Video Content Accessibility

Captions, summaries, and plain-language explanations can be added to educational, instructional, promotional, and consultation videos.

It can improve the accessibility of video platforms, online education, public information content, and corporate promotional content.

 

Customer Center and Consultation Automation

Consultations can be transcribed, results summarized, and follow-up actions organized.

When integrated with BoryTalk Solution, it can be extended with AI chatbot and consultation-automation capabilities.

 

Integrated Products and Solutions

 

BoryGlass Gov

It can be used to caption civil-service consultations at public institutions, hospitals, and welfare organizations, and to summarize consultations or convert them into plain language.

 

BoryGlass

It can display conversation captions on AR glasses and combine speech recognition with lip-reading STT to provide hearing-assistance services.

 

BoryVoice Solution

It can serve as a shared engine for speech recognition, lip-reading STT, real-time captions, video captions, and consultation records.

 

BORY Sense Solution

It can serve as the foundation platform for solutions that implement video understanding, captions, summaries, and plain-language conversion for a customer's environment.

 

BoryKiosk Solution

It can be used for kiosk voice ordering, captioned guidance, plain-language guidance, and UI support for older adults and digitally vulnerable users.

 

BoryTalk Solution

It can be used in AI chatbot and consultation automation for conversation input, consultation summaries, organization of customer questions, and generation of plain-language explanations.

 

Expected Benefits

 

Standardization of Accessibility Capabilities

Captioning, summarization, plain-language conversion, speech recognition, and lip-reading STT can be shared through a single platform rather than developed separately for each product.

 

Faster Expansion of Products and Solutions

With BorySense Platform, multiple products and solutions—including BoryGlass, BoryGlass Gov, BoryKiosk, BoryTalk, and BoryVoice—can be expanded more quickly.

 

Improved User Understanding

By providing audio, video, and conversation information as captions, summaries, and plain-language explanations, users can understand information more clearly.

 

Stronger Barrier-Free Service Competitiveness

For customers in public institutions, hospitals, educational institutions, and companies, it can offer value through enhanced accessibility, digital inclusion, support for vulnerable groups, and advanced hearing-assistance services.

 

More Efficient Consultation and Record-Keeping Work

It automatically transcribes and summarizes consultations and meetings, reducing the record-keeping workload for staff.

 

API-Based Business Expansion

BorySense APIs can be integrated with customers' existing websites, apps, kiosks, AR glasses, consultation systems, and video platforms to provide accessibility features.

 

Deployment Process

 

1. Requirements Review

We review the customer's service objectives, target users, usage environment, integrated systems, accessibility requirements, and security standards.

 

2. Data Environment Analysis

We analyze the audio-input environment, video quality, camera position, microphone configuration, lighting, noise, network, and data-retention policy.

 

3. Define the Functional Scope

We define the required capabilities among speech recognition, lip-reading STT, real-time captions, video captions, conversation summaries, plain-language conversion, accessibility UI, and API integration.

 

4. PoC Development

Using a limited environment or sample data, we validate caption quality, summary quality, plain-language conversion, screen display, and API integration feasibility.

 

5. Build the Platform

We build the AI engine, servers, database, dashboard, APIs, user screens, administrator screens, and security architecture for the customer's environment.

 

6. Integrate Products and Solutions

We integrate the required products and solutions, such as BoryGlass, BoryGlass Gov, BoryKiosk, BoryTalk, BoryVoice, and BORY Sense Solution.

 

7. Operational Verification and Enhancement

Based on actual usage data, we improve caption quality, recognition accuracy, summary results, plain-language expressions, screen readability, and user feedback.

 

What BorySense Platform Can Deliver

 

  • Build an AI Video Understanding Engine
  • Build a Speech Recognition STT Engine
  • Build a Lip-Reading STT Engine
  • Build an Audio-and-Lip Combined STT System
  • Real-Time Conversation Caption Generation
  • Generate captions for video content
  • Generate Public-Service Consultation Captions
  • Generate Hospital Reception Captions
  • AR Glasses Caption Display Integration
  • Integrate Kiosk Captioned Guidance
  • Summarize conversations
  • Automate Consultation Records
  • Convert text into plain language
  • Build accessible UI screens
  • Chatbot and Consultation System Integration
  • Integrate Meeting and Lecture Captioning Systems
  • External Service API Provision
  • Build an On-Premises Accessibility Engine
  • Build a Cloud-Based Accessibility API

 

Implementation Inquiries

BorySense Platform can be proposed as a cloud-based, on-premises, API-integrated, or product-and-solution-integrated configuration according to the customer's service objectives, data environment, accessibility requirements, security policies, and existing system architecture.

  • Inquiry About Implementing an AI Video Understanding and Conversational Accessibility Platform
  • Inquiry About Building a Real-Time Captioning Engine
  • Inquiry About Speech Recognition and Lip-Reading STT Engines
  • Inquiry About Building a Plain-Language Conversion Engine
  • Inquiry About Building a Conversation Summarization Engine
  • Inquiry About a Public Institution Civil-Service Accessibility Platform
  • Inquiry About a Hospital Reception and Consultation Accessibility Platform
  • Inquiry About an AR Glasses Caption API
  • Inquiry About a Kiosk Accessibility API
  • On-Premises Implementation Inquiry
  • Request a PoC Proposal

 

Important Information

BorySense Platform is an AI-based accessibility platform that analyzes video, audio, lip movements, and conversation data to provide captions, summaries, plain-language explanations, and accessible screens.

This platform assists communication and information understanding, but it does not guarantee complete recognition or interpretation of all audio, video, lip-movement, or conversation content.

The results of speech recognition, lip-reading STT, video understanding, summarization, and plain-language conversion may vary depending on ambient noise, lighting, camera position, microphone quality, speaking habits, video quality, network conditions, and data conditions.

In areas requiring important decisions, such as healthcare, legal, financial, or public-service matters, we recommend establishing a process in which responsible personnel review AI-generated captions, summaries, and plain-language explanations.

If consultation audio, video, captions, conversation records, or user information are stored or transmitted, prior consultation regarding privacy standards and data security policies is required.

This platform is intended to improve accessibility and assist service use; it should not be presented as a standalone system for medical diagnosis, legal judgment, or the creation of official evidence.

Before implementation, the intended use, data-collection method, caption-display method, analysis scope, storage scope, personal-data handling standards, integrated systems, and operating standards must be reviewed in advance.

 

Related Products, Solutions, and Platforms

 

Related Product

  • BoryGlass Gov / Lip-Reading STT Barrier-Free AI Civil-Service Consultation Package
  • BoryGlass / Bory Lip-Reading and Speech Recognition Hearing Assistance AR Glasses
  • BoryKiosk / Bory Barrier-Free AI Kiosk
  • BoryTalk / Bory AI Web Chatbot Implementation and AI Conversation Engine Package

 

Related Solution

  • BORY Sense Solution / Bory AI Video Understanding and Conversational Accessibility Solution
  • BoryVoice Solution / Bory Speech Recognition and Lip-Reading STT Solution
  • BoryGlass Gov Solution / Bory Barrier-Free AI Civil-Service Consultation Solution
  • BoryGlass Solution / Bory Lip-Reading and Speech Recognition AR Glasses Solution
  • BoryKiosk Solution / Bory Barrier-Free AI Customer Service, Ordering, and Payment Solution
  • BoryTalk Solution / Bory AI Chatbot and Consultation Automation Solution

 

Related Platform

  • BoryVoice Platform / Bory Voice and Conversational AI Platform
  • BoryVision Platform / Bory Video AI Platform
  • BoryXR Platform / Bory AR and XR Platform
  • BORY AI Platform / Bory Artificial Intelligence Platform
 



 


BORY.ai Key Services Overview

Bory Co., Ltd. develops products, solutions, platforms, and services for industrial, medical, public-sector, healthcare, and barrier-free applications based on AI technologies involving speech, language, video, sensors, and data.

Below are the main representative domains currently operated or being prepared by Bory Co., Ltd. 

 
Primary Domain Service/Brand Description
bory.ai BORY.ai The official AI brand website of Bory Co., Ltd., serving as the company’s main website for the integrated presentation of its products, solutions, platforms, and services
borysense.com BORY SENSE An AI-powered hearing assistance platform that supports communication for people with hearing disabilities and older adults through real-time captioning, lip-reading AI, and AR glasses integration
borytalk.com BORY TALK A web-based conversational AI chatbot service designed for civil service inquiries, consultations, information guidance, and customer support
borykiosk.com BORY KIOSK A barrier-free AI kiosk service for older adults and people with disabilities, featuring voice guidance, captioning, and easy-to-use interfaces
boryservice.com BORY SERVICE A service portal introducing Bory’s AI services and custom-built AI offerings for industrial, public-sector, and everyday applications
borysong.com BORY SONG A music AI service that supports AI-powered composition, music generation, and sound content production

 

 

 

  • 네이버 블러그 공유하기
  • 네이버 밴드에 공유하기
  • 페이스북 공유하기
  • 카카오스토리 공유하기