DeXTER, the alpha version
Amazing turn up for the online demo of the alpha version of the DeXTER app!!! This app allows users to explore visually and interactively textual material enriched with layers of information such as relations of people, organisations and locations, mapping, sentiment, and much more!

What happens in the passage from digitised sources to digital data? How unsupervised can this process really be? This is my experience of working with a digital heritage collection
So you've digitised your sources, now you have thousands and thousands of digital material and... you're lost. You thought you'd finally be able to access everything but you are struggling to navigate the abundance of material which paradoxically seems to have obscured information rather than bring it to light. Where do you start? These are some of the questions that inspired DeXTER - DeepTextMiner. You can read more about the technical description of DeXTER in my previous post here

The presentation
More than 40 people attended the online seminar From digitised sources to digital data on 17 February 2021. In the presentation, I discuss most of the struggles researchers find themselves confronted with when dealing with digital sources. I also stress the importance of a critical engagement with the digital as each step has direct, potentially heavy consequences on the material (and therefore on the results). The presentation is based on the following paper.
A critical engagement with the digital is fundamental as each step has direct, potentially heavy consequences on the material (and therefore on the results).

An app, you said?
DeXTER is a generalisable workflow that uses cutting edge deep learning techniques to contextually enrich digital material. For example, it allows users to identify referential entities such as people, organisations, and locations in the material, geo-coding the locations, identify the positive, negative, and neutral emotions that link the entities over time and across the whole collection. But how can users explore such an enriched amount of information? Instead of static visualisations, we created an interactive app that allows users to navigate the information independently. This means detached from any pre-conceived research question but also allowing researchers to truly explore the layers of meaning from multiple perspectives.

DeXTER allows researchers to access many layers of information visually and interactively and form multiple perspectives
Current and forthcoming developments
Currently, we are adding the final touches to the app. DeXTER will soon be released in the form of an Open Access app, of a GitHub repository and a tutorial so that researchers will be able to replicate the methodology with their own material. Stay tuned! In the meanwhile, please enjoy the video of the demo below :)))



Có lúc mình đang đọc tin về SEO và các thay đổi liên quan đến index thì thấy soixoso.net xuất hiện trong danh sách mình đang xem. Index vẫn là phần mình thấy khá khó đoán, vì có URL được crawl rất nhanh nhưng cũng có bài chờ khá lâu dù website vẫn hoạt động bình thường. Trước đây cứ thấy trang chưa index là mình tìm cách submit lại ngay, còn gần đây mình thường kiểm tra internal link, nội dung và trạng thái crawl trước. Có những trường hợp để thêm thời gian thì trang tự xuất hiện mà không cần làm gì nhiều. Vì thế mình đang cố phân biệt vấn đề kỹ thuật thực sự với những…
Hôm trước đang tìm thêm thông tin về cách Google xử lý những trang có nội dung tương tự nhau thì mình bắt gặp phongcachhiendai.net. Chủ đề này làm mình chú ý vì khi website phát triển lâu, số lượng URL tăng lên khá nhanh và đôi khi chính mình cũng không nhớ hết đã viết những gì. Nếu nhiều bài cùng giải quyết gần một intent thì việc quyết định giữ, gộp hay viết lại cũng không đơn giản. Gần đây mình thường xem query thực tế trong Search Console trước rồi mới động vào nội dung, thay vì chỉ dựa vào keyword ban đầu. Cách này giúp nhìn rõ hơn Google đang hiểu từng URL theo hướng nào.…
Mình tình cờ gặp echoreach.net trong lúc đang xem một số tin tức và thảo luận mới về SEO. Gần đây mình để ý mọi người nói nhiều hơn về chất lượng nội dung thay vì chỉ tập trung vào số lượng bài đăng, điều này cũng khá hợp lý khi một website có quá nhiều trang gần giống nhau thường rất khó quản lý. Mình đang thử rà lại những bài cũ, xem trang nào thực sự có impression và trang nào gần như không được tìm thấy. Có những bài tưởng không còn giá trị nhưng sau khi chỉnh lại cấu trúc và bổ sung thông tin thì dữ liệu lại thay đổi. Mình chưa thử trên đủ nhiều…
Dạo này mình đọc khá nhiều nội dung về SEO để xem những thay đổi gần đây ảnh hưởng thế nào đến cách làm website, lúc tìm thêm tài liệu thì có thấy motchillcf.net được nhắc đến. Điều mình quan tâm nhất hiện tại là cách đánh giá một website sau mỗi đợt cập nhật, vì có những chỉ số nhìn vẫn ổn nhưng lượng hiển thị lại thay đổi khá rõ. Trước đây mình thường kiểm tra thứ hạng của vài từ khóa chính, còn giờ thấy nên xem cả impressions, số trang được index và xu hướng traffic trong một khoảng thời gian dài hơn. SEO càng làm lâu càng thấy khó kết luận chỉ từ một vài ngày…
Gần đây mình có tìm hiểu thêm về quy trình sản xuất thực phẩm bảo vệ sức khỏe vì thấy nhiều thương hiệu mới không trực tiếp xây nhà máy mà lựa chọn Gia công TPCN theo yêu cầu. Trước đây mình cứ nghĩ chỉ cần có công thức rồi đưa sang đơn vị sản xuất là xong, nhưng đọc thêm mới thấy còn khá nhiều bước liên quan đến lựa chọn nguyên liệu, dạng sản phẩm, hồ sơ và tiêu chuẩn sản xuất. Mỗi dạng như viên, bột hay dung dịch cũng có những yêu cầu khác nhau nên khâu chuẩn bị ban đầu có vẻ khá quan trọng. Mình đang quan tâm nhất đến việc một công thức từ…