Occitan OCR in C# and .NET
IronOCR은 .NET 개발자가 오크어를 포함한 126개 언어로 된 이미지와 PDF 문서에서 텍스트를 읽을 수 있도록 해주는 C# 소프트웨어 구성 요소입니다. 이는 .NET 개발자 전용으로 개발된 Tesseract의 고급 포크 버전으로, 속도와 정확도 면에서 다른 Tesseract 엔진보다 뛰어난 성능을 보여줍니다.
IronOcr.언어.오크어의 내용
이 패키지에는 .NET용 OCR 언어 46개가 포함되어 있습니다.
- 오크어
- 오크시탄베스트
- 오크시탄패스트
다운로드
오크어 언어 팩
설치
먼저 Occitan OCR 패키지를 .NET 프로젝트에 설치해야 합니다.
코드 예제
이 C# 코드 예제는 이미지나 PDF 문서에서 오크어 텍스트를 읽어옵니다.
// Importing the IronOCR namespace
using IronOcr;
class Program
{
static void Main()
{
// Create a new instance of the OCR engine for Occitan language
var Ocr = new IronTesseract();
Ocr.Language = OcrLanguage.Occitan;
// Use a using block for proper disposal of resources
using (var Input = new OcrInput(@"images\Occitan.png"))
{
// Perform OCR on the input image
var Result = Ocr.Read(Input);
// Retrieve the recognized text
var AllText = Result.Text;
// Output the recognized text to the console
Console.WriteLine(AllText);
}
}
}' Importing the IronOCR namespace
Imports IronOcr
Module Program
Sub Main()
' Create a new instance of the OCR engine for Occitan language
Dim Ocr As New IronTesseract()
Ocr.Language = OcrLanguage.Occitan
' Use a Using block for proper disposal of resources
Using Input As New OcrInput("images\Occitan.png")
' Perform OCR on the input image
Dim Result = Ocr.Read(Input)
' Retrieve the recognized text
Dim AllText = Result.Text
' Output the recognized text to the console
Console.WriteLine(AllText)
End Using
End Sub
End Module이 예제는 오크어 텍스트가 포함된 이미지 파일에서 텍스트를 읽도록 IronOCR 라이브러리를 구성하는 방법을 보여줍니다. 이 프로그램은 OCR 언어를 오크어(Occitan)로 설정하고 이미지를 처리하여 인식된 텍스트를 출력합니다.

커티스 차우는 칼턴 대학교에서 컴퓨터 과학 학사 학위를 취득했으며, Node.js, TypeScript, JavaScript, React를 전문으로 하는 프론트엔드 개발자입니다. 직관적이고 미적으로 뛰어난 사용자 인터페이스를 만드는 데 열정을 가진 그는 최신 프레임워크를 활용하고, 잘 구성되고 시각적으로 매력적인 매뉴얼을 제작하는 것을 즐깁니다.