TAILIEUCHUNG - Báo cáo khoa học: "Unsupervised Word Alignment with Arbitrary Features"

We introduce a discriminatively trained, globally normalized, log-linear variant of the lexical translation models proposed by Brown et al. (1993). In our model, arbitrary, nonindependent features may be freely incorporated, thereby overcoming the inherent limitation of generative models, which require that features be sensitive to the conditional independencies of the generative process. However, unlike previous work on discriminative modeling of word alignment (which also permits the use of arbitrary features), the parameters in our models are learned from unannotated parallel sentences, rather than from supervised word alignments. . | Unsupervised Word Alignment with Arbitrary Features Chris Dyer Jonathan Clark Alon Lavie Noah A. Smith Language Technologies Institute Carnegie Mellon University PittsbUrgh PA 15213 UsA cdyer jhclark alavie nasmith @ Abstract We introduce a discriminatively trained globally normalized log-linear variant of the lexical translation models proposed by Brown et al. 1993 . In our model arbitrary nonindependent features may be freely incorporated thereby overcoming the inherent limitation of generative models which require that features be sensitive to the conditional independencies of the generative process. However unlike previous work on discriminative modeling of word alignment which also permits the use of arbitrary features the parameters in our models are learned from unannotated parallel sentences rather than from supervised word alignments. Using a variety of intrinsic and extrinsic measures including translation performance we show our model yields better alignments than generative baselines in a number of language pairs. 1 Introduction Word alignment is an important subtask in statistical machine translation which is typically solved in one of two ways. The more common approach uses a generative translation model that relates bilingual string pairs using a latent alignment variable to designate which source words or phrases generate which target words. The parameters in these models can be learned straightforwardly from parallel sentences using EM and standard inference techniques can recover most probable alignments Brown et al. 1993 . This approach is attractive because it only requires parallel training data. An alternative to the generative approach uses a discriminatively trained 409 alignment model to predict word alignments in the parallel corpus. Discriminative models are attractive because they can incorporate arbitrary overlapping features meaning that errors observed in the predictions made by the model can be addressed by engineering new

Lan Vy 44 11 pdf

Upload

Bấm vào đây để xem trước nội dung

Tải xuống

TÀI LIỆU LIÊN QUAN

Báo cáo khoa học: "Smaller Alignment Models for Better Translations: Unsupervised Word Alignment with the 0"

9 73 0

Báo cáo khoa học: "Fully Unsupervised Word Segmentation with BVE and MDL"

6 43 0

Báo cáo khoa học: "Bayesian Unsupervised Word Segmentation with Nested Pitman-Yor Language Modeling"

9 36 0

Báo cáo khoa học: "Unsupervised Word Alignment with Arbitrary Features"

11 37 0

Báo cáo khoa học: "An Algorithm for Unsupervised Transliteration Mining with an Application to Word Alignment"

10 50 0

Báo cáo khoa học: "Efﬁcient Unsupervised Discovery of Word Categories Using Symmetric Patterns and High Frequency Words"

8 45 0

Báo cáo khoa học: "Contextual Dependencies in Unsupervised Word Segmentation∗"

8 59 0

Báo cáo khoa học: "Unsupervised Learning of Acoustic Sub-word Units"

4 70 0

Báo cáo khoa học: "UNSUPERVISED WORD SENSE DISAMBIGUATION RIVALING SUPERVISED METHODS"

8 53 0

Báo cáo khoa học: "Combining Unsupervised Lexical Knowledge Methods for Word Sense Disambiguation"

8 58 0

TÀI LIỆU XEM NHIỀU

Một Case Về Hematology (1)

8 461863 55

Giới thiệu :Lập trình mã nguồn mở

14 22634 59

Tiểu luận: Tư tưởng Hồ Chí Minh về xây dựng nhà nước trong sạch vững mạnh

13 10884 529

Câu hỏi và đáp án bài tập tình huống Quản trị học

14 10064 446

Phân tích và làm rõ ý kiến sau: “Bài thơ Tự tình II vừa nói lên bi kịch duyên phận vừa cho thấy khát vọng sống, khát vọng hạnh phúc của Hồ Xuân Hương”

3 9518 104

Ebook Facts and Figures – Basic reading practice: Phần 1 – Đặng Tuấn Anh (Dịch)

249 8278 1125

Tiểu luận: Nội dung tư tưởng Hồ Chí Minh về đạo đức

16 8230 423

Mẫu đơn thông tin ứng viên ngân hàng VIB

8 7864 2220

Đề tài: Dự án kinh doanh thời trang quần áo nữ

17 6674 253

Vật lý hạt cơ bản (1)

29 5769 85

TỪ KHÓA LIÊN QUAN

TÀI LIỆU MỚI ĐĂNG

Sáng tạo trong thuật toán và lập trình với ngôn ngữ Pascal và C# Tập 2 - Chương 4

47 246 1 26-04-2024

BeginningMac OS X Tiger Dashboard Widget Development 2006 phần 2

34 211 0 26-04-2024

Trading Strategies Profit Making Techniques For Stock_8

23 175 0 26-04-2024

MySQL Basics for Visual Learners PHẦN 9

15 183 0 26-04-2024

Công nghiệp gang thép Việt Nam : Một giai đoạn phát triển và chuyển đổi chính sách mới part 5

6 194 0 26-04-2024

Lịch sử Đội TNTP Hồ Chí Minh - CHƯƠNG III VÂNG LỜI BÁC DẠY, LÀM NGHÌN VIỆC TỐT, CHỐNG MỸ, CỨU NƯỚC, THIẾU NIÊN SĂN SÀNG

45 136 0 26-04-2024

The profit magic of stock Timing The Markets_5

22 119 0 26-04-2024

Khurana et al. Journal of Orthopaedic Surgery and Research 2010, 5:23

7 133 0 26-04-2024

báo cáo hóa học:" Endoscopic decompression for intraforaminal and extraforaminal nerve root compression"

7 107 0 26-04-2024

Diseases of the Liver and Biliary System - part 1

33 123 0 26-04-2024

TÀI LIỆU HOT

Mẫu đơn thông tin ứng viên ngân hàng VIB

8 7864 2220

Giáo trình Tư tưởng Hồ Chí Minh - Mạch Quang Thắng (Dành cho bậc ĐH - Không chuyên ngành Lý luận chính trị)

152 5718 1364

Ebook Chào con ba mẹ đã sẵn sàng

112 3767 1231

Ebook Tuyển tập đề bài và bài văn nghị luận xã hội: Phần 1

62 5318 1136

Ebook Facts and Figures – Basic reading practice: Phần 1 – Đặng Tuấn Anh (Dịch)

249 8278 1125

Giáo trình Văn hóa kinh doanh - PGS.TS. Dương Thị Liễu

561 3498 643

Tiểu luận: Tư tưởng Hồ Chí Minh về xây dựng nhà nước trong sạch vững mạnh

13 10884 529

Giáo trình Sinh lí học trẻ em: Phần 1 - TS Lê Thanh Vân

122 3683 525

Giáo trình Pháp luật đại cương: Phần 1 - NXB ĐH Sư Phạm

274 4045 514

Bài tập nhóm quản lý dự án: Dự án xây dựng quán cafe

35 4127 480