วันพุธที่ 19 มิถุนายน พ.ศ. 2567

การหาหัวข้อวิจัยเป็นขั้นตอนสำคัญที่อย่าให้ใครมาทำแทน

การที่อาจารย์มี list หัวข้อวิจัยไว้ให้นักศึกษาเลือกshopเป็นแนวปฏิบัติที่ไม่ดี เพราะการหาหัวข้อวิจัยเกิดจากกระบวนการทบทวนวรรณกรรมและวิเคราะห์สังเคราะห์จนมองเห็นปัญหา ทักษะเหล่านี้จะไม่เกิดขึ้นในตัวผู้เรียนถ้าไปช็อปหัวข้อวิจัยมาทำเลย นอกจากนี้หัวข้อวิจัยที่จัดเตรียมไว้แล้วมักมีแนวทางแก้ปัญหาไว้อยู่ด้วยทำให้ผู้เรียนแทบไม่ต้องคิดหาวิธีแก้ปัญหาด้วยตัวเองยิ่งทำให้ผู้เรียนขาดทักษะในการแก้ปัญหา หัวข้อวิจัยสำเร็จรูปก็เปรียบเหมือนบะหมี่กึ่งสำเร็จรูปชงกินได้เร็วแต่ไม่ดีต่อร่างกาย หรืออาจเปรียบเหมือนสอนทำเมนูปลาแต่ไม่ได้สอนการจับปลาด้วยตัวเองทำให้ได้ทักษะความรู้ไปไม่ครบloop กลยุทธ์แบบกลยุทธ์แบบนี้เน้นปริมาณผู้เรียนแต่ไม่เน้นคุณภาพ

สิ่งที่ควรทำมากกว่าคือกำหนดพื้นที่การทำวิจัยให้ผู้เรียนเลือกอาจารย์ที่ปรึกษาได้เหมาะสมกับความสนใจของตนเองและทราบ specific field ที่ตัวเองสนใจในพื้นที่การทำวิจัยดังกล่าวสำหรับเข้าไปดำเนินการทบทวนวรรณกรรมต่อไป

วันพฤหัสบดีที่ 6 มิถุนายน พ.ศ. 2567

Lagrangian relaxation

 https://en.wikipedia.org/wiki/Lagrangian_relaxation

วันอาทิตย์ที่ 2 มิถุนายน พ.ศ. 2567

Multidiscipline vs Interdiscipline

  • Multidisciplinarity draws on knowledge from different disciplines but stays within their boundaries. 
  • Interdisciplinarity analyzes, synthesizes and harmonizes links between disciplines into a coordinated and coherent whole.

วันศุกร์ที่ 31 พฤษภาคม พ.ศ. 2567

Article/Publication types

ACM:

  • Research Article
  • Short Paper e.g. letter 
  • Review Article
  • Survey Article
  • Technical Note
  • Tutorial
  • Interview
  • Note
IEEE

  • Journals & Magazines
  • Conference Proceedings
    • Abstract: Synopsis of your research (250 words or less)
    • Extended abstract: High-level summary of your research (less than 2 pages)
    • Brief or short paper: Summary of your research (less than 4 pages)
    • Full paper: Complete paper describing your research in full (6-8 pages)
  • Books

Roboflow

PaaS for building computer vision applications end‑to‑end—from data collection and annotation to model training and deployment.

To annotate images, build, and deploy computer vision models.

https://roboflow.com/

วันพฤหัสบดีที่ 30 พฤษภาคม พ.ศ. 2567

วันเสาร์ที่ 11 พฤษภาคม พ.ศ. 2567

Cloudflare's captcha (Turnstile)

With Turnstile, we adapt the actual challenge outcome to the individual visitor or browser. First, we run a series of small non-interactive JavaScript challenges gathering more signals about the visitor/browser environment. Those challenges include, proof-of-work, proof-of-space, probing for web APIs, and various other challenges for detecting browser-quirks and human behavior. As a result, we can fine-tune the difficulty of the challenge to the specific request and avoid ever showing a visual puzzle to a user. สรุป Turnstile ส่งสคริปมาเก็บข้อมูลสภาพแวดล้อมในเครื่องเพื่อดูว่าเป็น web browser จริงหรือเป็น bot

cf. https://www.cloudflare.com/en-gb/products/turnstile/

https://developers.cloudflare.com/turnstile/

วันอังคารที่ 7 พฤษภาคม พ.ศ. 2567

Generative AI

It is capable of generating text, images, videos, or other data.

In other words, traditional AI excels at pattern recognition, while generative AI excels at pattern creation. Traditional AI can analyze data and tell you what it sees, but generative AI can use that same data to create something entirely new.

AI search engine

It's actually a RAG LLM.


https://www.perplexity.ai/


https://devhub.in.th/blog/perplexity-ai-search

วันศุกร์ที่ 26 เมษายน พ.ศ. 2567

วันพุธที่ 24 เมษายน พ.ศ. 2567

Well-known Thai hosted conferences

IC2IT & NCCIT ปิดรับบทความกลาง มค และกลาง มีค ตามลำดับ

ECTI-CON ปิดรับบทความสิ้น มค

JCSSE ปิดรับบทความต้น เมย

INCIT & NCIT ปิดรับบทความสิ้น สค และกลาง กย ตามลำดับ

วันจันทร์ที่ 22 เมษายน พ.ศ. 2567

Widely used hashing algorithm

The most commonly used hashing algorithm is SHA-256 (Secure Hash Algorithm 256). It is widely used for cryptographic security and data integrity verification in various applications, including digital signatures, SSL/TLS certificates, and blockchain technology.

Vulnerabilities of MD5 (hash value is 128 bits) 

1. Collision Vulnerabilities: MD5 is susceptible to collision attacks, where two different inputs produce the same hash. This poses a severe security risk, particularly in applications like digital signatures.

2. Preimage Attacks: Attackers can reverse-engineer the hash to find an input that matches a given MD5 hash, compromising data security.

3. Speed: The speed at which MD5 can generate hashes makes it susceptible to brute force attacks.

cf. https://medium.com/@techclaw/exploring-the-power-and-vulnerabilities-of-the-md5-algorithm-feb249ef9dfb#:~:text=MD5%20is%20susceptible%20to%20collision,in%20applications%20like%20digital%20signatures.

Colmap

ใช้สร้างโมเดล 3 มิติจากภาพถ่ายเพื่อไปหมุนดูบนจอหรือในแว่นตา VR

https://colmap.github.io/

วันอาทิตย์ที่ 21 เมษายน พ.ศ. 2567

Percentile vs Quartile in Academic publication

  • เลข percentile ยิ่งมากยิ่งดี เพราะเป็นพื้นที่ใต้กราฟ normal distribution เริ่มจากซ้ายสุดของกราฟไปทางขวา ในการจัดอันดับวารสารและรายงานประชุมวิชาการ เลข percentile ยิ่งเยอะยิ่งดี
    • Top 10% หรือ Tier 1 Journal : ค่าเปอร์เซ็นไทล์ 90 ขึ้นไป (ตรวจสอบจาก Scopus ค้นด้วยชื่อวารสาร รายละเอียด Source details จะแสดง Percentile ของวารสารในฟิลด์ที่เกี่ยวข้องต่างๆ)
    • Q1Journal : ค่าเปอร์เซ็นไทล์ 75-99
    • Q2Journal : ค่าเปอร์เซ็นไทล์ 50-74
    • Q3Journal : ค่าเปอร์เซ็นไทล์ 25-49
    • Q4Journal : ค่าเปอร์เซ็นไทล์ 0-24
  • Q1 = 25th percentile, Q2 = 50th percentile, Q3 = 75th percentile, Q4 = 100th percentile จริงๆเลข Quartile ยิ่งเยอะยิ่งดีเพราะได้คะแนนเยอะ แต่ในการจัดคุณภาพวารสารเลข Quartile ยิ่งน้อยกลับยิ่งดี
===

The standard formula used to find the value at a specific percentile in a sorted data set is:
L=(P/100)(N+1)
In this formula, L represents the position (index) of the value in the sorted list, P is the desired percentile, and N is the total number of data points in your data set.
1. Arrange Data
Sort your data set in ascending order (from lowest to highest).
2. Calculate the Position
Plug your desired percentile P and total number of scores N into the formula to find the position L. For example, to find the 90th percentile (P = 90) of 19 test scores (N = 19):
L=(90/100)(19+1)=18
The 18th value in your sorted list is the 90th percentile.

วันอังคารที่ 9 เมษายน พ.ศ. 2567

Low code development platform

Outsystem 

https://www.stream.co.th/develop-application-by-outsystems/

https://youtu.be/8FXHSZaln6U?si=S9Yndsrc-nH8nLZ2

Opposite is high code.

Information entropy

Information entropy is a measure of uncertainty in an event.  The higher the entropy, the less predictable the information and the more "surprised" you'd be by the outcome. It's like a measure of how much information you actually need in order to convey the sample space of the event, on average. By average, it means the "expected value" (see my post Average vs Expected value) or expected amount of information you need to encode the sample space of event. (The expected value is a form of average value but calculated from probability rather than straigntforwardly dividing by N.)

Formula 

H(X) คือระดับความไม่แน่นอนของเหตุการณ์ X ซึ่งเป็นการทอยเหรียญ 1 เหรียญ และค่าที่เป็นไปได้มี 2 ค่าคือ {หัว,ก้อย} นี่คือ X ใต้ sigma; p(x) คือความน่าจะเป็นที่จะเกิดเหตุการณ์ x

ในการทอยเหรียญ 1 เหรียญ ความน่าจะเป็นในการออกหัวคือ 0.7 ออกก้อยคือ 0.3 จึงคำนวณ H(X) ได้เป็น -(0.7log_2(0.7)+0.3log_2(0.3)) = 0.8816 นี่คือความไม่แน่นอนเฉลี่ยของการทอยเหรียญอันนี้ หรือกล่าวอีกนัยหนึ่งต้องใช้บิทจำนวน 0.8816 บิท (2^0.8816) ในการ encode เหตุการณ์นี้ซึ่งจะออกเป็นหัว 0.7 ก้อย 0.3  แต่ถ้าเป็นเหรียญที่ fair คือโอกาสออกหัวและก้อยเท่ากันคือ 0.5 จะได้ H(X) =1 คือใช้ 1 บิท (2^1 = 2 ค่าที่อาจเกิดขึ้นในการทอย)

Average VS Expected value

  •  Expected Value is used in case of Random Variables (or in other words Probability Distributions). Since, the average is defined as the sum of all the elements divided by the sum of their frequencies. But for the case of Probability distribution we can't describe a random variable in terms of its frequency beforehand, thus we use the probability instead. Conceptually, probability of an element is frequency of an event divided by size of sample space (N). Thus, the average in case of random variable can be given by sum of probabilities multiplied by its respective event (where p(x)*x is conceptually frequency of x divided by total frequency).
  • Average on the other hand is used in case where we have the knowledge of frequencies of individual elements and total count of the elements, for example, in case of known data set or sample. We can simply use the fundamental definition of average to calculate it.

  • สรุปว่า Average คือค่าเฉลี่ยที่คำนวณเมื่อรู้ N และจำนวน (frequency) ที่ชัดเจนของแต่ละค่าเหตุการณ์ เช่นมีคนได้คะแนน 85 สองคน (freq=2) 84 หนึ่งคน (freq=1) 82 ห้าคน (freq=5) ... ส่วน Expected value คือค่าเฉลี่ยที่คำนวณจากความน่าจะเป็นของแต่ละค่าเหตุการณ์ (random variable) เพราะไม่รู้จำนวนที่ชัดเจน

  • Example

Lottery Game

You pay $5 to play.

Possible prizes:

 
$10 (probability = 0.2),
$20 (probability = 0.1),
$50 (probability = 0.05),
$0 (probability = 0.65). 
Net winnings () คือรางวัลที่หักต้นทุน 5$ ออก:

 (if you win $10),

 (if you win $20),

 (if you win $50),

X
=
5
X=−5 (if you win $0).
E[X]=(5⋅0.2)+(15⋅0.1)+(45⋅0.05)+(−5⋅0.65) =1+1.5+2.25−3.25=1.5